Text Retrieval from Natural and Scanned Images
Jayshree Ghorpade-Aher, Sumeet Gajbhar, Amey Sarode, Govardhan Gayake, Piyush Daund · International Journal of Computer Applications · 2016
Digital documents are easy to handle, share and store than hard copy of documents.These made people to prefer digital document over hard copy of documents.Digital documents are nothing but scanned images of a document or natural images of notice boards, traffic signs.Text detection is an important process required to extract text from images.Text from images can be extracted using Optical Character Recognition (OCR).OCR works in three phases as preprocessing, segmentation, character recognition.Preprocessing is the first phase which uses different techniques for making text easy to extract from images.In segmentation phase, each character is isolated.Then this will be given as input to OCR recognition phase which will compare it with training data-set and will recognize character.In this survey paper, different techniques for OCR are discussed.