Designing and Implementing of the Webpage Information Extracting Model Based on Tags

Xu Zhang, Dong Feng Yan · 2011

In this article, a novel model of Webpage information extraction based on tags is presented. With the ingenious algorithm, the model preformed better than Html Parser and Jsoup in most cases. It can be a URL filter of the Net Crawler in order to enhance efficiency.

Read the paper · More papers on PaperTik