Study on HowNet-Based Word Similarity Algorithm

Xiaofeng Gu · Zhongwen xinxi xuebao · 2010

Word(sentence) similarity computing based on the HowNet usually treats the optimal matches between the primitives or words as the basic unit,and the ultimate outcome can be the sum of weighted counts.However,this approach often results in the information duplication and some irrational constructions.To deal with these issues,this paper propose to calculate the similarity of sets by the statistics on common information(commonality) and the different information(differences) between the two sets of direct primitives.Moreover,the paper introduces this measure into the calculation of sentence similarity.The final experimental analysis shows that the proposed method is more stable and effective.

Read the paper · More papers on PaperTik