Measurement of word similarity based on Corpus
Shao Xiao-min · Journal of Computer Applications · 2006
A semantic relevant database named Corpus was built to store the required information in word similarity measurement. Corpus got the information from large scale text training and store the information in word space and relation space after analysis and tailoring. The word similarity measurement algorithm by constructing the context relation vectors based on Corpus was given, which proved to be a feasible method by experiments.