Pixel and semantic capabilities from an image-object based document representation

Michael J. Gormish, Kathrin Berkner, Martin P. Boliek, Guotong Feng, Edward L. Schwartz · Proceedings of SPIE, the International Society for Optical Engineering/Proceedings of SPIE · 2007

This paper reports on novel and traditional pixel and semantic operations using a recently standardized document representation called JPM. The JPM representation uses compressed pixel arrays for all visible elements on a page. Separate data containers called boxes provide the layout and additional semantic information. JPM and related image-based document representation standards were designed to obtain the most rate efficient document compression. The authors, however, use this representation directly for operations other than compression typically performed either on pixel arrays or semantic forms. This paper describes the image representation used in the JPM standard and presents techniques to (1) perform traditional raster-based document analysis on the compressed data, (2) transmit semantically meaningful portions of compressed data between devices, (3) create multiple views from one compressed data stream, and (4) edit high resolution document images with only low resolution proxy images.

Read the paper · More papers on PaperTik