Automatic discovery of attribute words from Web documents
Kosuke Tokunaga, Jun’ichi Kazama, Kentaro Torisawa · Institutional Repositories DataBase (IRDB) · 2005
Abstract. We propose a method of acquiring attribute words for a wide range of objects from Japanese Web documents. The method is a simple unsupervised method that utilizes the statistics of words, lexico-syntactic patterns, and HTML tags. To evaluate the attribute words, we also es-tablish criteria and a procedure based on question-answerability about the candidate word. 1