An Algorithm for Community Identification and Dynamical Addition Based on Web Pages Contents Similarity and Link Relation

Wang Chuan-bao · Journal of Zhengzhou University · 2011

An algorithm for community identification based on the Web pages contents similarity and the link relation between the Web pages was proposed.The algorithm not only considered the hyperlinks between Web pages but focused on the content similarity of Web pages.This method overcame the limitations of ignoring the content of Web pages in traditional community discovery algorithms,so that the communities founded in the content were more relevant.In addition,the paper added the new members based on the original community dynamically,and added the new Web pages which linked to the Web pages of original community related to the theme into the original community.Experiments showed that the method was applied to community discovery in the network,and the community was more relevant in the content.

Read the paper · More papers on PaperTik