DocumentCode
1376859
Title
Enhanced speech separation in room acoustic environments with selected binaural cues
Author
Cho, Namgook ; Kuo, C. C Jay
Author_Institution
Univ. of Southern California, Los Angeles, CA, USA
Volume
55
Issue
4
fYear
2009
fDate
11/1/2009 12:00:00 AM
Firstpage
2163
Lastpage
2171
Abstract
We propose a robust technique to separate audio sources received by a microphone array in a room acoustic environment with an underdetermined mixing process (i.e., the number of sources is larger than the number of mixtures). Our scheme consists of two stages: 1) estimation of mixing parameters and 2) recovery of source signals. For the first stage, contrary to the traditional DUET-type methods that exploit all binaural cues, we estimate the mixing parameters by selecting a reliable subset of binaural cues based on the phase determinacy condition and source sparsity. As a result, we can determine the mixing parameters successfully even in a reverberant environment with longer time delay. Then, proper mathematical tools are applied to the underdetermined linear system to recover the original audio sources for the second stage. Experimental results on simulated data in a room acoustic environment are given to show a significant gain over the DUET-type method in audio source separation.
Keywords
architectural acoustics; microphone arrays; source separation; speech recognition; audio sources; binaural cues; enhanced speech separation; microphone array; phase determinacy condition; room acoustic environments; source sparsity; Acoustic noise; Automatic speech recognition; Delay effects; Delay estimation; Loudspeakers; Microphone arrays; Parameter estimation; Source separation; Speech enhancement; Terrorism; Audio source separation, underdetermined mixing, room acoustics, sparse representation, multichannel audio;
fLanguage
English
Journal_Title
Consumer Electronics, IEEE Transactions on
Publisher
ieee
ISSN
0098-3063
Type
jour
DOI
10.1109/TCE.2009.5373783
Filename
5373783
Link To Document