Extracting the characters from scanned Telugu document images

N. Subramanyam, Kishore Surendra, M. Hanumanthu · ZENITH International Journal of Multidisciplinary Research · 2014

Segmentation is an important task of any OCR system. It separates the text document images into lines, words and primitives (characters). The accuracy of OCR system mainly depends on the segmentation algorithm being used. Segmentation Telugu text is difficult when compared with Latin based languages because of its structural complexity and increased character set. It contains vowels, consonants and compound characters.

Read the paper · More papers on PaperTik