Comparison of Approaches for Information Extraction from the Web

Michal Toman · 2008

In this paper we compare two new methods for information extraction from the web pages. The first method is based on statistical analysis of web page content and the second one uses the XQueries for information extraction from semi- structured documents. We compare precision and recall rate of both automatic methods with manually created extracts. The paper includes a comparison of our extraction methods with other methods used for information extraction from web sources.

Read the paper · More papers on PaperTik