• DocumentCode
    3429701
  • Title

    Models of English text

  • Author

    Teahan, W.J. ; Cleary, John G.

  • Author_Institution
    Dept. of Comput. Sci., Waikato Univ., Hamilton, New Zealand
  • fYear
    1997
  • fDate
    25-27 Mar 1997
  • Firstpage
    12
  • Lastpage
    21
  • Abstract
    The problem of constructing models of English text is considered. A number of applications of such models including cryptology, spelling correction and speech recognition are reviewed. The best current models for English text have been the result of research into compression. Not only is this an important application of such models but the amount of compression provides a measure of how well such models perform. Three main classes of models are considered: character based models, word based models, and models which use auxiliary information in the form of parts of speech. These models are compared in terms of their memory usage and compression
  • Keywords
    cryptography; data compression; natural languages; speech recognition; spelling aids; word processing; English text models; auxiliary information; character based models; cryptology; data compression; memory usage; speech recognition; spelling correction; word based models; Application software; Computer science; Cryptography; Ice; Natural languages; Performance evaluation; Predictive models; Speech recognition; Text recognition;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Data Compression Conference, 1997. DCC '97. Proceedings
  • Conference_Location
    Snowbird, UT
  • ISSN
    1068-0314
  • Print_ISBN
    0-8186-7761-9
  • Type

    conf

  • DOI
    10.1109/DCC.1997.581953
  • Filename
    581953