Model-based control strategy for document image analysis

Frank Fein, Frank Hoenes · Proceedings of SPIE, the International Society for Optical Engineering/Proceedings of SPIE · 1992

Generally, document analysis and understanding involves many processing steps, like unskewing, segmentation, logical labeling, text recognition, and text analysis. Most of these steps can be subdivided into different tasks depending on the problem-solving methods available. All of the techniques are more or less specialized to certain input, but some are also competitive. As a consequence, a document analysis system incorporating many analysis methods must properly schedule and control these methods to obtain an optimal result. In this paper, we present a model for the control strategy of a document image analysis system as well as mechanisms for its interpretation that describe three important aspects: which specialist can be applied to which object in which analysis state. The analysis model comprises all possible sequences of processing steps which are relevant for the analysis tasks. The underlying document architecture supports the analysis specialists by corresponding knowledge and provides a framework for representing the analysis results.

Read the paper · More papers on PaperTik