Similarity Computing of Documents Based on Weighted Semantic Network

Binbin Yang · Journal of Intelligence · 2012

The traditional documents similarity algorithm based on the thought of statistical information of word frequency only considers the w eight of feature items in a document,thus ignores the semantic relations among feature items.This paper considers both the importance of feature items in a document and the semantic relations among feature items,and proposes to construct a w eighted semantic netw ork of document feature items to calculate the similarity of documents.In the process of constructing the model,there are some appropriate improvements in the selection of feature items and the calculation of feature items w eight.With an experiment,it is w ell-proved that,compared w ith the traditional algorithm,the suggested algorithm based on w eighted semantic netw ork promotes the precision of the calculation of documents similarity.

Read the paper · More papers on PaperTik