期刊论文详细信息
International Journal of Computer Science and Security
Morphological Reconstruction for Word Level Script Identification.
Mallikarjun Hangarge1  B. V. Dhandra1 
[1] $$
关键词: Script identification;    Bilingual documents;    OCR;    Morphological reconstruction;    regional descriptors;   
DOI  :  
来源: Computer Science and Security
PDF
【 摘 要 】

A line of a bilingual document page may contain text words in regional languageand numerals in English. For Optical Character Recognition (OCR) of such adocument page, it is necessary to identify different script forms before running anindividual OCR system. In this paper, we have identified a tool of morphologicalopening by reconstruction of an image in different directions and regionaldescriptors for script identification at word level, based on the observation thatevery text has a distinct visual appearance. The proposed system is developedfor three Indian major bilingual documents, Kannada, Telugu and Devnagaricontaining English numerals. The nearest neighbour and k-nearest neighbouralgorithms are applied to classify new word images. The proposed algorithm istested on 2625 words with various font styles and sizes. The results obtained arequite encouraging

【 授权许可】

Unknown   

【 预 览 】
附件列表
Files Size Format View
RO201912040511409ZK.pdf 164KB PDF download
  文献评价指标  
  下载次数:6次 浏览次数:16次