Designing and Implementing of the Webpage Information Extracting Model Based on Tags
Xu Zhang, Dong Feng Yan · 2011
In this article, a novel model of Webpage information extraction based on tags is presented. With the ingenious algorithm, the model preformed better than Html Parser and Jsoup in most cases. It can be a URL filter of the Net Crawler in order to enhance efficiency.