Title :
A system for the automatic layout segmentation and classification of digital documents
Author :
Cinque, L. ; Levialdi, S. ; Malizia, A.
Author_Institution :
Dept. of Inf. Sci., Rome Univ., Italy
Abstract :
Paper document recognition is fundamental for office automation becoming every day a more powerful tool in those fields where information is still on paper. Document recognition follows from data acquisition, from both journals and entire books, in order to transform them into digital objects. We present a new system for document recognition that follows the open source methodologies, XML description for document segmentation and classification, which turns out to be beneficial in terms of classification precision, and general-purpose availability.
Keywords :
XML; document image processing; image classification; image segmentation; object recognition; office automation; XML; automatic layout segmentation; data acquisition; digital document classification; digital objects; document analysis; document segmentation; office automation; open source methodologies; paper document recognition; Data acquisition; Data preprocessing; Image analysis; Image processing; Image recognition; Image segmentation; Merging; Pattern analysis; Text analysis; XML;
Conference_Titel :
Image Analysis and Processing, 2003.Proceedings. 12th International Conference on
Print_ISBN :
0-7695-1948-2
DOI :
10.1109/ICIAP.2003.1234050