DocumentCode
2105293
Title
An auditory feature extraction method for robust speaker recognition
Author
Fengsong Hu ; Xiaoyu Cao
Author_Institution
Coll. of Inf. Sci. & Eng., Hunan Univ., Changsha, China
fYear
2012
fDate
9-11 Nov. 2012
Firstpage
1067
Lastpage
1071
Abstract
Based on simulate human auditory model, the paper proposed an auditory cepstrum coefficient for speaker recognition. This paper analyzed the working mechanism of the human auditory model, and simulated the auditory model of human ear cochlea by gammatone filter banks instead of the traditional triangular filter banks; Based on the nonlinear signal processing capability of human auditory model, exponential compression is used instead of the fixed logarithm compression used in the classic MFCC. Finally, the simulation experiments were conducted based on GMM recognition algorithm. The experimental results show that the proposed auditory cepstrum coefficient has better noise robustness than MFCC and LPCC.
Keywords
Gaussian processes; channel bank filters; feature extraction; speaker recognition; GMM recognition algorithm; LPCC; MFCC; auditory cepstrum coefficient; auditory feature extraction method; fixed logarithm compression; gammatone filter banks; human auditory model; human ear cochlea model; nonlinear signal processing capability; robust speaker recognition; triangular filter banks; Auditory model; Feature extraction; Gammatone Filter; Speaker recognition;
fLanguage
English
Publisher
ieee
Conference_Titel
Communication Technology (ICCT), 2012 IEEE 14th International Conference on
Conference_Location
Chengdu
Print_ISBN
978-1-4673-2100-6
Type
conf
DOI
10.1109/ICCT.2012.6511354
Filename
6511354
Link To Document