Fast and Large-scale Unsupervised Relation Extraction
Sho Takase, Naoaki Okazaki, Naoaki Okazaki, 80227, 50601118, Kentaro Inui, 80364, 60272689 · Institutional Repositories DataBase (IRDB) · 2015
A common approach to unsupervised relation extraction builds clusters of patterns expressing the same relation.In order to obtain clusters of relational patterns of good quality, we have two major challenges: the semantic representation of relational patterns and the scalability to large data.In this paper, we explore various methods for modeling the meaning of a pattern and for computing the similarity of patterns mined from huge data.In order to achieve this goal, we apply algorithms for approximate frequency counting and efficient dimension reduction to unsupervised relation extraction.The experimental results show that approximate frequency counting and dimension reduction not only speeds up similarity computation but also improves the quality of pattern vectors.