A Similarity Algorithm for Chinese Text Based on Semantics
Xia Zhimin · Computer and Modernization · 2015
This paper computes the semantic similarity of words using the How Net and extracting the text keywords to compute the similarity of the texts. After segmenting the text and filtering stop words,it calculates the weights of word to extract the key words of the text by combining the gender,word frequency and paragraph frequency of the word. By calculating the similarity of the keywords,the similarity value of the texts is calculated. The analysis of the significant difference of the experimental results shows that its accuracy is further improved compared with the traditional semantic algorithm and vector space model algorithm.