• DocumentCode
    3296509
  • Title

    An Improved Template-Based Approach to Keyword Spotting Applied to the Spoken Content of User Generated Video Blogs

  • Author

    Barakat, M.S. ; Ritz, C.H. ; Stirling, D.A.

  • Author_Institution
    ICT Res. Inst., Univ. of Wollongong, Wollongong, NSW, Australia
  • fYear
    2012
  • fDate
    9-13 July 2012
  • Firstpage
    723
  • Lastpage
    728
  • Abstract
    This paper presents a new technique for preparing word templates to improve the performance of dynamic time warping based keyword spotting. The proposed technique selects one reference template from a small set of examples and in contrast to existing model based approaches does not require extensive training. Precision and recall results from applying the technique to template selection for use in searching for keywords in a clean speech database and within a set of user generated video blogs are superior to existing approaches used to select a template. As opposed to automatic speech recognition approaches, the technique is promising for use in searching for keywords that are not adequately represented in training databases.
  • Keywords
    Web sites; video retrieval; word processing; DTW; KWS; clean speech database; dynamic time warping-based keyword spotting; keyword searching; performance improvement; precision value; recall value; reference template; spoken contents; template selection; training databases; user-generated video blogs; word template-based approach; Blogs; Clustering algorithms; Databases; Feature extraction; Hidden Markov models; Speech; Vectors; Dynamic Time Warping (DTW); Hidden Markov Model (HMM); K-Medoid; Social Networks; Spotting (KWS); Users Video Blogs;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Multimedia and Expo (ICME), 2012 IEEE International Conference on
  • Conference_Location
    Melbourne, VIC
  • ISSN
    1945-7871
  • Print_ISBN
    978-1-4673-1659-0
  • Type

    conf

  • DOI
    10.1109/ICME.2012.10
  • Filename
    6298488