DocumentCode :
542314
Title :
Implementation of automatic capitalisation generation systems for speech input
Author :
Kim, Ji-Hwan ; Woodland, P.C.
Author_Institution :
Cambridge University Engineering Department, Trumpington Street, CB2 IPZ, United Kingdom
Volume :
1
fYear :
2002
fDate :
13-17 May 2002
Abstract :
In this paper, two systems are proposed for the task of capitalisation generation. The first system is a slightly modified speech recogniser. In this system, every word in the vocabulary is duplicated: once in a decapitalised form and again in capitalised forms. In addition, the language model is re-trained on mixed case texts. The other system is based on Named Entity (NE) recognition and punctuation generation, since most capitalised words are first words in sentences or NE words. Both systems are compared for speech input. The system based on NE recognition and punctuation generation shows better results in Word Error Rate (WER) and in F-measure than the system modified from the speech recogniser.
Keywords :
Benchmark testing; Data models; Logic gates; Vocabulary; World Wide Web;
fLanguage :
English
Publisher :
ieee
Conference_Titel :
Acoustics, Speech, and Signal Processing (ICASSP), 2002 IEEE International Conference on
Conference_Location :
Orlando, FL, USA
ISSN :
1520-6149
Print_ISBN :
0-7803-7402-9
Type :
conf
DOI :
10.1109/ICASSP.2002.5743874
Filename :
5743874
Link To Document :
بازگشت