DocumentCode
1299845
Title
Speaker Diarization Based on Intensity Channel Contribution
Author
Barra-Chicote, Roberto ; Pardo, Jose Manuel ; Ferreiros, Javier ; Montero, Juan Manuel
Author_Institution
Speech Technol. Group, Univ. Politec. de Madrid, Madrid, Spain
Volume
19
Issue
4
fYear
2011
fDate
5/1/2011 12:00:00 AM
Firstpage
754
Lastpage
761
Abstract
The time delay of arrival (TDOA) between multiple microphones has been used since 2006 as a source of information (localization) to complement the spectral features for speaker diarization. In this paper, we propose a new localization feature, the intensity channel contribution (ICC) based on the relative energy of the signal arriving at each channel compared to the sum of the energy of all the channels. We have demonstrated that by joining the ICC features and the TDOA features, the robustness of the localization features is improved and that the diarization error rate (DER) of the complete system (using localization and spectral features) has been reduced. By using this new localization feature, we have been able to achieve a 5.2% DER relative improvement in our development data, a 3.6% DER relative improvement in the RT07 evaluation data and a 7.9% DER relative improvement in the last year´s RT09 evaluation data.
Keywords
feature extraction; microphone arrays; speaker recognition; spectral analysis; time-of-arrival estimation; DER; ICC feature; TDOA; diarization error rate; intensity channel contribution; localization feature; multiple microphone; speaker diarization; speaker segmentation; spectral feature; time delay of arrival; Intensity channel contribution (ICC); speaker diarization; speaker segmentation; speech processing in meetings;
fLanguage
English
Journal_Title
Audio, Speech, and Language Processing, IEEE Transactions on
Publisher
ieee
ISSN
1558-7916
Type
jour
DOI
10.1109/TASL.2010.2062507
Filename
5551177
Link To Document