Extracting Domain-Relevant Term Using Wikipedia Based on Random Walk Model

Wenjuan Wu, Tao Liu, He Hu, Xiaoyong Du · 2012

In this paper we present a new approach for the automatic identification of domain-relevant concepts and entities of a given domain using the category and page structures of the Wikipedia in a language independent way. By applying Markov random walk algorithm on the weighted Wikipedia link graph, our approach can identify large quantities of domain-relevant concepts and entities with very little human effort. Experimental results show that our method achieves high accuracy and acceptable efficiency in domain-relevant term extraction.

Read the paper · More papers on PaperTik