Study on HowNet-Based Word Similarity Algorithm
Xiaofeng Gu · Zhongwen xinxi xuebao · 2010
Word(sentence) similarity computing based on the HowNet usually treats the optimal matches between the primitives or words as the basic unit,and the ultimate outcome can be the sum of weighted counts.However,this approach often results in the information duplication and some irrational constructions.To deal with these issues,this paper propose to calculate the similarity of sets by the statistics on common information(commonality) and the different information(differences) between the two sets of direct primitives.Moreover,the paper introduces this measure into the calculation of sentence similarity.The final experimental analysis shows that the proposed method is more stable and effective.