International Journal of Computer Science and Security | |
Morphological Reconstruction for Word Level Script Identification. | |
Mallikarjun Hangarge1  B. V. Dhandra1  | |
[1] $$ | |
关键词: Script identification; Bilingual documents; OCR; Morphological reconstruction; regional descriptors; | |
DOI : | |
来源: Computer Science and Security | |
【 摘 要 】
A line of a bilingual document page may contain text words in regional languageand numerals in English. For Optical Character Recognition (OCR) of such adocument page, it is necessary to identify different script forms before running anindividual OCR system. In this paper, we have identified a tool of morphologicalopening by reconstruction of an image in different directions and regionaldescriptors for script identification at word level, based on the observation thatevery text has a distinct visual appearance. The proposed system is developedfor three Indian major bilingual documents, Kannada, Telugu and Devnagaricontaining English numerals. The nearest neighbour and k-nearest neighbouralgorithms are applied to classify new word images. The proposed algorithm istested on 2625 words with various font styles and sizes. The results obtained arequite encouraging
【 授权许可】
Unknown
【 预 览 】
Files | Size | Format | View |
---|---|---|---|
RO201912040511409ZK.pdf | 164KB | download |