Title of article :
Bounded support Gaussian mixture modeling of speech spectra
Author/Authors :
J.، Lindblom, نويسنده , , J.، Samuelsson, نويسنده ,
Issue Information :
روزنامه با شماره پیاپی سال 2003
Pages :
-87
From page :
88
To page :
0
Abstract :
Lately, Gaussian mixture (GM) models have found new applications in speech processing, and particularly in speech coding. This paper provides a review of GM based quantization and prediction. The main contribution is a discussion on GM model optimization. Two previously presented algorithms of EM-type are analyzed in some detail, and models are estimated and evaluated experimentally using theoretical measures as well as GM based speech spectrum coding and prediction. It has been argued that since many sources have a bounded support, this should be utilized in both the choice of model, and the optimization algorithm. By low-dimensional modeling examples, illustrating the behavior of the two algorithms graphically, and by full-scale evaluation of GM based systems, the advantages of a bounded support approach are quantified. For all evaluation techniques in the study, model accuracy is improved when the bounded support approach is adopted. The gains are typically largest for models with diagonal covariance matrices.
Keywords :
Laminated waveguide , low-temperature co-fired ceramic (LTCC) , rectangular waveguide (RWG) , millimeter wave , waveguide transition
Journal title :
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
Serial Year :
2003
Journal title :
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
Record number :
86890
Link To Document :
بازگشت