• DocumentCode
    2100473
  • Title

    A system for the automatic layout segmentation and classification of digital documents

  • Author

    Cinque, L. ; Levialdi, S. ; Malizia, A.

  • Author_Institution
    Dept. of Inf. Sci., Rome Univ., Italy
  • fYear
    2003
  • fDate
    17-19 Sept. 2003
  • Firstpage
    201
  • Lastpage
    206
  • Abstract
    Paper document recognition is fundamental for office automation becoming every day a more powerful tool in those fields where information is still on paper. Document recognition follows from data acquisition, from both journals and entire books, in order to transform them into digital objects. We present a new system for document recognition that follows the open source methodologies, XML description for document segmentation and classification, which turns out to be beneficial in terms of classification precision, and general-purpose availability.
  • Keywords
    XML; document image processing; image classification; image segmentation; object recognition; office automation; XML; automatic layout segmentation; data acquisition; digital document classification; digital objects; document analysis; document segmentation; office automation; open source methodologies; paper document recognition; Data acquisition; Data preprocessing; Image analysis; Image processing; Image recognition; Image segmentation; Merging; Pattern analysis; Text analysis; XML;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Image Analysis and Processing, 2003.Proceedings. 12th International Conference on
  • Print_ISBN
    0-7695-1948-2
  • Type

    conf

  • DOI
    10.1109/ICIAP.2003.1234050
  • Filename
    1234050