• DocumentCode
    2972845
  • Title

    The Asian network-based speech-to-speech translation system

  • Author

    Sakti, Sakriani ; Kimura, Noriyuki ; Paul, Michael ; Hori, Chiori ; Sumita, Eiichiro ; Nakamura, Satoshi ; Park, Jun ; Wutiwiwatchai, Chai ; Xu, Bo ; Riza, Hammam ; Arora, Karunesh ; Luong, Chi Mai ; Li, Haizhou

  • Author_Institution
    Nat. Inst. of Inf. & Commun. Technol. (NICT), Japan
  • fYear
    2009
  • fDate
    Nov. 13 2009-Dec. 17 2009
  • Firstpage
    507
  • Lastpage
    512
  • Abstract
    This paper outlines the first Asian network-based speech-to-speech translation system developed by the Asian Speech Translation Advanced Research (A-STAR) consortium. The system was designed to translate common spoken utterances of travel conversations from a certain source language into multiple target languages in order to facilitate multiparty travel conversations between people speaking different Asian languages. Each A-STAR member contributes one or more of the following spoken language technologies: automatic speech recognition, machine translation, and text-to-speech through Web servers. Currently, the system has successfully covered 9 languages-namely, 8 Asian languages (Hindi, Indonesian, Japanese, Korean, Malay, Thai, Vietnamese, Chinese) and additionally, the English language. The system´s domain covers about 20,000 travel expressions, including proper nouns that are names of famous places or attractions in Asian countries. In this paper, we discuss the difficulties involved in connecting various different spoken language translation systems through Web servers. We also present speech-translation results on the first A-STAR demo experiments carried out in July 2009.
  • Keywords
    language translation; speech recognition; Asian languages; Asian network; Asian speech translation advanced research; automatic speech recognition; machine translation; speech-to-speech translation; text-to-speech; Automatic speech recognition; Automation; Communications technology; Computer networks; Information technology; Natural languages; Paper technology; Speech synthesis; Telecommunication computing; Web server;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Automatic Speech Recognition & Understanding, 2009. ASRU 2009. IEEE Workshop on
  • Conference_Location
    Merano
  • Print_ISBN
    978-1-4244-5478-5
  • Electronic_ISBN
    978-1-4244-5479-2
  • Type

    conf

  • DOI
    10.1109/ASRU.2009.5373353
  • Filename
    5373353