Calculation of Chinese Words Semantic Similarity Using Network Search Engines
Gao Guo-qian · Computer Technology and Development · 2014
Similarity computation of Chinese words is a key problem in Chinese information processing. It measures semantic similarity between Chinese words using the information returned by web search engines. First,implement a model named WebPMI which computes similarity using page counts,and then,describe another model named CODC which analyzes semantic similarity using text snippets. Finally,present the algorithm based on the two models. Experimental results show that this algorithm outperforms all the existing web- based semantic similarity measures for Chinese,and is close to the traditional semantic similarity measures using lexicon.