Text Localization and Detection Method for Born-digital Images

Samabia Tehsin, Asif Masood, Sumaira Kausar, Younus Javed · IETE Journal of Research · 2013

AbstractMultimedia data has increased rapidly in recent years. Textual information present in multimedia contains important information about the image/video content. The proposed method provides very efficient way to extract text from Born-Digital images. Firstly, edges are extracted from a grayscale image. New edge detection technique is introduced in this research, which gives better results for low-contrast web images. Then morphological operators are applied on the image. These operators are used to connect the broken edges of objects present in an image. Each object is classified as text or non-text on the basis of text features such as size, height to width ratio, and binary transitions, using K-Means clustering. Two new features, namely horizontal fluctuation count and vertical fluctuation count, are introduced in the proposed work. Dataset of International Conference on Document Analysis and Recognition 2011 Robust Reading Competition, Challenge 1: “Reading Text in Born-Digital Images (Web and Em...

Read the paper · More papers on PaperTik