Research on Automatic Retrieval and Classification for Chinese RSS Information
Jin Zhang · Jisuanji gongcheng · 2011
This paper presents a web crawler fitting for RSS which uses breadth-first algorithm and focuses on RSS to carry out automatically collection.And based on word segment,it improves the method to calculate word weight,works on word filtering,and implements automatically classification aiming at RSS using VSM.Experimental result shows that the system achieves to retrieve and classify Chinese RSS information with lower system cost and higher accuracy.And it can take manage of RSS information syndication effectively.