A clustering algorithm for Chinese adjectives and nouns

Yang Wen, Chunfa Yuan, Changning Huang · 2000

This paper proposes a bidirctional hierarchical clustering algorithm for simultaneously clustering words of different parts of speech based on collocations. The algorithm is composed of cycles of two kinds of alternate clustering processes. We construct an objective function based on Minimum Description Length. To partly solve the problem caused by sparse data two concepts of collocational degree and revisional distance are presented.

Read the paper · More papers on PaperTik