• Title of article

    Efficient decoding strategies for conversational speech recognition using a constrained nonlinear statespace model

  • Author/Authors

    Deng، Kung-Li نويسنده , , J.Z.، Ma, نويسنده ,

  • Issue Information
    روزنامه با شماره پیاپی سال 2003
  • Pages
    -58
  • From page
    59
  • To page
    0
  • Abstract
    In this paper, we present two efficient strategies for likelihood computation and decoding in a continuous speech recognizer using an underlying nonlinear state-space dynamic model for the hidden speech dynamics. The statespace model has been specially constructed so as to be suitable for the conversational or casual style of speech where phonetic reduction abounds. Two specific decoding algorithms, based on optimal state-sequence estimation for the nonlinear state-space model, are derived, implemented, and evaluated. They successfully overcome the exponential growth in the original search paths by using the path-merging approaches derived from Bayesʹ rule. We have tested and compared the two algorithms using the speech data from the Switchboard corpus, confirming their effectiveness. Conversational speech recognition experiments using the Switchboard corpus further demonstrated that the use of the new decoding strategies is capable of reducing the recognizerʹs word error rate compared with two baseline recognizers, including the HMM system and the nonlinear state-space model using the HMM-produced phonetic boundaries, under identical test conditions.
  • Keywords
    Laminated waveguide , millimeter wave , rectangular waveguide (RWG) , waveguide transition , low-temperature co-fired ceramic (LTCC)
  • Journal title
    IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
  • Serial Year
    2003
  • Journal title
    IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
  • Record number

    86934