Header menu link for other important links
Syntactic and Semantic labeling of hierarchically organized document image components of Indian scripts
Published in
Pages: 314 - 317
In this paper we describe our document image analysis system which performs segmentation, content characterization as well as semantic labeling of components. Segmentation is done using white spaces and gives the segmented components arranged in a hierarchy. Semantic labeling is done using domain knowledge which is specified where possible in the form of a document model applicable to a class of documents. The novelty of the system lies in the suite of methods it employs which are capable of handling documents in Indian scripts. We have obtained promising results for semantic segmentation of over 30 categories of documents in Indian scripts. © 2009 IEEE.
About the journal
JournalProceedings of the 7th International Conference on Advances in Pattern Recognition, ICAPR 2009