A clustering algorithm for Chinese adjectives and nouns
Yang Wen, Chunfa Yuan, Changning Huang · 2000
This paper proposes a bidirctional hierarchical clustering algorithm for simultaneously clustering words of different parts of speech based on collocations. The algorithm is composed of cycles of two kinds of alternate clustering processes. We construct an objective function based on Minimum Description Length. To partly solve the problem caused by sparse data two concepts of collocational degree and revisional distance are presented.