DocumentCode
3510998
Title
Gender independent Bangla automatic speech recognition
Author
Hassan, F. ; Kotwal, M.R.A. ; Khan, M.S.A. ; Huda, M.N.
Author_Institution
Dept. of Comput. Sci. & Eng., United Int. Univ., Dhaka, Bangladesh
fYear
2012
fDate
18-19 May 2012
Firstpage
144
Lastpage
148
Abstract
Speaker-specific characteristics play an important role on the performance of Bangla (widely used as Bengali) automatic speech recognition (ASR). Gender factor shows adverse effect in the classifier while recognizing a speech by an opposite gender, such as, training a classifier by male but testing is done by female or vice-versa. To obtain a robust ASR system in practice it is necessary to invent a system that incorporates gender independent effect for particular gender. In this paper, we have proposed a Gender-Independent technique for ASR that focused on a gender factor. The proposed method trains the classifier with the both types of gender, male and female, and evaluates the classifier for the male and female. For the experiments, we have designed a medium size Bangla (widely known as Bengali) speech corpus for both the male and female. The proposed system has showed a significant improvement of word correct rates, word accuracies and sentence correct rates in comparison with the method that suffers from gender effects. Moreover, it requires a fewer mixture component in hidden Markov model (HMMs) and hence, computation time.
Keywords
gender issues; hidden Markov models; natural language processing; signal classification; speaker recognition; ASR system; Bangla speech corpus; Bengali; HMM; classifier training; female classifier; gender factor; gender independent Bangla automatic speech recognition; hidden Markov model; sentence correct rates; speaker-specific characteristics; word accuracy; word correct rates; Dentistry; Electronic publishing; Hidden Markov models; Ice; Information services; Internet; Mel frequency cepstral coefficient; automatic speech recognition; gender factor; hidden Markov model; sentence correct rates; word accuracies; word correct rates;
fLanguage
English
Publisher
ieee
Conference_Titel
Informatics, Electronics & Vision (ICIEV), 2012 International Conference on
Conference_Location
Dhaka
Print_ISBN
978-1-4673-1153-3
Type
conf
DOI
10.1109/ICIEV.2012.6317500
Filename
6317500
Link To Document