A novel approach for improving recognition accuracies in OCR of printed telugu text

Chellapilla Vasantha Lakshmi, C. Patvardhan, Mohit Prasad · 2005

Telugu is one of the oldest and popular languages of India spoken by more than 66 million people especially in South India. Not much work has been reported on the development of optical character recognition systems for Telugu text. Therefore, it is an area of current research. During the process of recognition, it is observed that, in many cases, a symbol is recognized erroneously because the recognizer incorrectly outputs a very similar looking symbol. Several such sets of symbols that are commonly confused for each other are identified and presented in a table called confusion table. Special logic and algorithms are developed using simple structural features, for resolving confusion and improving recognition accuracies considerably without too much additional computational effort.

Read the paper · More papers on PaperTik