Title :
Implementation of automatic capitalisation generation systems for speech input
Author :
Kim, Ji-Hwan ; Woodland, P.C.
Author_Institution :
Cambridge University Engineering Department, Trumpington Street, CB2 IPZ, United Kingdom
Abstract :
In this paper, two systems are proposed for the task of capitalisation generation. The first system is a slightly modified speech recogniser. In this system, every word in the vocabulary is duplicated: once in a decapitalised form and again in capitalised forms. In addition, the language model is re-trained on mixed case texts. The other system is based on Named Entity (NE) recognition and punctuation generation, since most capitalised words are first words in sentences or NE words. Both systems are compared for speech input. The system based on NE recognition and punctuation generation shows better results in Word Error Rate (WER) and in F-measure than the system modified from the speech recogniser.
Keywords :
Benchmark testing; Data models; Logic gates; Vocabulary; World Wide Web;
Conference_Titel :
Acoustics, Speech, and Signal Processing (ICASSP), 2002 IEEE International Conference on
Conference_Location :
Orlando, FL, USA
Print_ISBN :
0-7803-7402-9
DOI :
10.1109/ICASSP.2002.5743874