Picture, graphics, and text classification of document image regions

Shriram V. Revankar, Zhigang Fan · Proceedings of SPIE, the International Society for Optical Engineering/Proceedings of SPIE · 2000

Various rendering techniques are being used for document reproduction and printing. Some rendering techniques work better for text, some for graphics and some others work better for picture regions. Therefore, dividing a document image into regions that need to be rendered differently from its neighboring regions is useful for good reproduction and printing. In this paper we describe a method to classify previously segmented regions of a page image into three classes, namely text, graphics and pictures. In addition to printing and copying, this classification of regions into broad basic classes is also useful for automatic storage and retrieval, and efficient communication of document images.

Read the paper · More papers on PaperTik