Skip navigation
Please use this identifier to cite or link to this item: http://repository.iitr.ac.in/handle/123456789/5660
Title: Multioriented and curved text lines extraction from Indian documents
Authors: Pal U.
Pratim Roy, Partha
Published in: IEEE Transactions on Systems, Man, and Cybernetics, Part B: Cybernetics
Abstract: There are printed artistic documents where text lines of a single page may not be parallel to each other. These text lines may have different orientations or the text lines may be curved shapes. For the optical character recognition (OCR) of these documents, we need to extract such lines properly. In this paper, we propose a novel scheme, mainly based on the concept of water reservoir analogy, to extract individual text lines from printed Indian documents containing multioriented and/or curve text lines. A reservoir is a metaphor to illustrate the cavity region of a character where water can be stored. In the proposed scheme, at first, connected components are labeled and identified either as isolated or touching. Next, each touching component is classified either straight type (S-type) or curve type (C-type), depending on the reservoir base-area and envelope points of the component. Based on the type (S-type or C-type) of a component two candidate points are computed from each touching component. Finally, candidate regions (neighborhoods of the candidate points) of the candidate points of each component are detected and after analyzing these candidate regions, components are grouped to get individual text lines. © 2004 IEEE.
Citation: IEEE Transactions on Systems, Man, and Cybernetics, Part B: Cybernetics (2004), 34(4): 1676-1684
URI: https://doi.org/10.1109/TSMCB.2004.827613
http://repository.iitr.ac.in/handle/123456789/5660
Issue Date: 2004
ISSN: 10834419
Author Scopus IDs: 57200742116
56880478500
Author Affiliations: Pal, U., Comp. Vis./Pattern Recognition Unit, Indian Statistical Institute, Kolkata-108, India
Roy, P.P., Tata Consultancy Service, Kolkata-91, India
Corresponding Author: Pal, U.; Comp. Vis./Pattern Recognition Unit, Indian Statistical Institute, Kolkata-108, India; email: umapada@isical.ac.in
Appears in Collections:Journal Publications [CS]

Files in This Item:
There are no files associated with this item.
Show full item record


Items in Repository are protected by copyright, with all rights reserved, unless otherwise indicated.