Document Image Analysis via Model Checking
Marco Aiello · UvA-DARE (University of Amsterdam) · 2002
Introduction When Dave placed his own drawing in front of the `eye' of HAL---in 2001: A Space Odyssey---HAL showed to have correctly comprehended and interpreted the sketch. "That's Dr. Hunter, isn't it?" [9]. But what would have happened if Dave used the first page of a newspaper in front of the eye and started discussing its contents? Considering HAL a system capable of AI, we expect HAL to recognize the document as a newspaper, to understand how to extract information and to understand its contents. Finally, we expect Dave and HAL to begin a conversation on the contents of the document. Here we present a methodology based on model checking, which has been successfully experimented on an heterogeneous collection of documents [1, 11], to extract the content from images of documents. We focus on mechanically generated documents, in contrast with hand-writing and sketches. Using terms better-known to the image processing community, we are interested in logical structure detection in t