Speaker Diarization Based on Intensity Channel Contribution


Autoria(s): Barra Chicote, Roberto; Pardo Muñoz, José Manuel; Ferreiros López, Javier; Montero Martínez, Juan Manuel
Data(s)

2011

Resumo

The time delay of arrival (TDOA) between multiple microphones has been used since 2006 as a source of information (localization) to complement the spectral features for speaker diarization. In this paper, we propose a new localization feature, the intensity channel contribution (ICC) based on the relative energy of the signal arriving at each channel compared to the sum of the energy of all the channels. We have demonstrated that by joining the ICC features and the TDOA features, the robustness of the localization features is improved and that the diarization error rate (DER) of the complete system (using localization and spectral features) has been reduced. By using this new localization feature, we have been able to achieve a 5.2% DER relative improvement in our development data, a 3.6% DER relative improvement in the RT07 evaluation data and a 7.9% DER relative improvement in the last year's RT09 evaluation data.

Formato

application/pdf

Identificador

http://oa.upm.es/11778/

Idioma(s)

eng

Publicador

E.T.S.I. Telecomunicación (UPM)

Relação

http://oa.upm.es/11778/2/INVE_MEM_2011_106872.pdf

http://ieeexplore.ieee.org/xpls/abs_all.jsp?arnumber=5551177

info:eu-repo/semantics/altIdentifier/doi/10.1109/TASL.2010.2062507

Direitos

http://creativecommons.org/licenses/by-nc-nd/3.0/es/

info:eu-repo/semantics/openAccess

Fonte

IEEE Transactions on Audio, Speech and Language Processing, ISSN 1558-7916, 2011, Vol. 19, No. 4

Palavras-Chave #Telecomunicaciones #Electrónica
Tipo

info:eu-repo/semantics/article

Artículo

PeerReviewed