Title :
Gaussian Specific Compensation for Channel Distortion in Speech Recognition
Author :
He, Yongjun ; Han, Jiqing
Author_Institution :
Sch. of Comput. Sci. & Technol., Harbin Inst. of Technol., Harbin, China
Abstract :
Channel distortion is one of the major factors degrading the performance of automatic speech recognition (ASR) systems. Most of the current compensation methods rely on the assumption that the channel distortion remains unchanged within an utterance or globally. However, we show in this letter that the distortion varies over speech frames even if the channel response is unchanged. To address this problem, we relax the above-mentioned assumption and propose a new method to compensate the channel distortion for each Gaussian of the acoustic models. Firstly, we derive the relationship between the clean and distorted models, and then estimate the channel magnitude response with the expectation-maximization (EM) algorithm. Finally, we obtain the matched models with the estimated magnitude response and the clean models. Experiments were conducted on the TIMIT/NTIMIT databases and the results confirmed the effectiveness of the proposed method.
Keywords :
Gaussian processes; expectation-maximisation algorithm; speech recognition; Gaussian specific compensation; TIMIT-NTIMIT databases; acoustic models; automatic speech recognition systems; channel distortion; channel magnitude response estmation; channel response; expectation-maximization algorithm; Acoustic distortion; Databases; Hidden Markov models; Mel frequency cepstral coefficient; Speech; Speech recognition; Automatic speech recognition; channel distortion; expectation-maximization;
Journal_Title :
Signal Processing Letters, IEEE
DOI :
10.1109/LSP.2011.2165058