DocumentCode :
1253882
Title :
Speechbot: an experimental speech-based search engine for multimedia content on the web
Author :
Van Thong, Jean-Manuel ; Moreno, Pedro J. ; Logan, Beth ; Fidler, Blair ; Maffey, Katrina ; Moores, Matthew
Author_Institution :
Compaq Comput. Corp., Cambridge, MA, USA
Volume :
4
Issue :
1
fYear :
2002
fDate :
3/1/2002 12:00:00 AM
Firstpage :
88
Lastpage :
96
Abstract :
As the Web transforms from a text-only medium into a more multimedia-rich medium, the need arises to perform searches based on the multimedia content. In this paper, we present an audio and video search engine to tackle this problem. The engine uses speech recognition technology to index spoken audio and video files from the World Wide Web (WWW) when no transcriptions are available. If transcriptions (even imperfect ones) are available, we can also take advantage of them to improve the indexing process. Our engine indexes several thousand talk and news radio shows covering a wide range of topics and speaking styles from a selection of public Web sites with multimedia archives. Our Web site is similar in spirit to normal Web search sites; it contains an index, not the actual multimedia content. The audio from these shows suffers in acoustic quality due to bandwidth limitations, coding, compression, and poor acoustic conditions. Our word error rate (WER) results using appropriately trained acoustic models show remarkable resilience to the high compression, although many factors combine to increase the average WERs over standard broadcast news benchmarks. We show that, even if the transcription is inaccurate, we can still achieve good retrieval performance for typical user queries (77.5%)
Keywords :
database indexing; multimedia databases; search engines; speech recognition; Speechbot; Web search; multimedia content; search engine; speech recognition; word error rate; Bandwidth; Error analysis; Indexing; Multimedia communication; Resilience; Search engines; Speech recognition; Web search; Web sites; World Wide Web;
fLanguage :
English
Journal_Title :
Multimedia, IEEE Transactions on
Publisher :
ieee
ISSN :
1520-9210
Type :
jour
DOI :
10.1109/6046.985557
Filename :
985557
Link To Document :
بازگشت