Graph-based Word Clustering Considering the Distance and the Connectivity of a Co-occurrence

Supaporn Simcharoen, Herwig Unger · 2022

Word clustering is a typical method of natural language processing. Several approaches for word clustering have been developed which consider different factors. The following article presents two factors, including the closest distance and the connectivity of a co-occurrence. The classical clustering algorithms, including k-means and Chinese Whispers, are chosen to compare their cluster quality. The results show that the quality of both proposed clustering algorithms of these two factors is close to k-means clustering.

Read the paper · More papers on PaperTik