• DocumentCode
    2770531
  • Title

    Speechfind for CDP: Advances in spoken document retrieval for the U. S. collaborative digitization program

  • Author

    Kim, Wooil ; Hansen, John H L

  • Author_Institution
    Univ. of Texas at Dallas, Richardson
  • fYear
    2007
  • fDate
    9-13 Dec. 2007
  • Firstpage
    687
  • Lastpage
    692
  • Abstract
    This paper presents our recent advances for SpeechFind, a CRSS-UTD designed spoken document retrieval system for the U.S. based Collaborative Digitization Program (CDP). A proto-type of SpeechFind for the CDP is currently serving as the search engine for 1,300 hours of CDP audio content which contain a wide range of acoustic conditions, vocabulary and period selection, and topics. In an effort to determine the amount of user corrected transcripts needed to impact automatic speech recognition (ASR) and audio search, a web-based online interface for verification of ASR-generated transcripts was developed. The procedure for enhancing the transcription performance for SpeechFind is also presented. A selection of adaptation methods for language and acoustic models are employed depending on the acoustics of the corpora under test. Experimental results on the CDP corpus demonstrate that the employed model adaptation scheme using the verified transcripts is effective in improving recognition accuracy. Through a combination of feature/acoustic model enhancement and language model selection, up to 24.8% relative improvement in ASR was obtained. The SpeechFind system, employing automatic transcript generation, online CDP transcript correction, and our transcript reliability estimator, demonstrates a comprehensive support mechanism to ensure reliable transcription and search for U.S. libraries with limited speech technology experience.
  • Keywords
    Internet; audio acoustics; indexing; online front-ends; search engines; speech processing; speech recognition; speech-based user interfaces; vocabulary; ASR-generated transcript verification; CDP audio content; CRSS-UTD designed spoken document retrieval system; SpeechFind system; US based Collaborative Digitization Program; Web-based online interface; acoustic conditions; audio indexing; audio search; automatic speech recognition; period selection; search engine; topics; vocabulary; Acoustic testing; Adaptation model; Automatic speech recognition; Collaboration; Information retrieval; Libraries; Material storage; Search engines; Speech recognition; Vocabulary; CDP; NGSW; SpeechFind; model enhancement; spoken document retrieval; transcript verification;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Automatic Speech Recognition & Understanding, 2007. ASRU. IEEE Workshop on
  • Conference_Location
    Kyoto
  • Print_ISBN
    978-1-4244-1746-9
  • Electronic_ISBN
    978-1-4244-1746-9
  • Type

    conf

  • DOI
    10.1109/ASRU.2007.4430195
  • Filename
    4430195