DocumentCode
2100473
Title
A system for the automatic layout segmentation and classification of digital documents
Author
Cinque, L. ; Levialdi, S. ; Malizia, A.
Author_Institution
Dept. of Inf. Sci., Rome Univ., Italy
fYear
2003
fDate
17-19 Sept. 2003
Firstpage
201
Lastpage
206
Abstract
Paper document recognition is fundamental for office automation becoming every day a more powerful tool in those fields where information is still on paper. Document recognition follows from data acquisition, from both journals and entire books, in order to transform them into digital objects. We present a new system for document recognition that follows the open source methodologies, XML description for document segmentation and classification, which turns out to be beneficial in terms of classification precision, and general-purpose availability.
Keywords
XML; document image processing; image classification; image segmentation; object recognition; office automation; XML; automatic layout segmentation; data acquisition; digital document classification; digital objects; document analysis; document segmentation; office automation; open source methodologies; paper document recognition; Data acquisition; Data preprocessing; Image analysis; Image processing; Image recognition; Image segmentation; Merging; Pattern analysis; Text analysis; XML;
fLanguage
English
Publisher
ieee
Conference_Titel
Image Analysis and Processing, 2003.Proceedings. 12th International Conference on
Print_ISBN
0-7695-1948-2
Type
conf
DOI
10.1109/ICIAP.2003.1234050
Filename
1234050
Link To Document