Graph-based Word Clustering Considering the Distance and the Connectivity of a Co-occurrence
Supaporn Simcharoen, Herwig Unger · 2022
Word clustering is a typical method of natural language processing. Several approaches for word clustering have been developed which consider different factors. The following article presents two factors, including the closest distance and the connectivity of a co-occurrence. The classical clustering algorithms, including k-means and Chinese Whispers, are chosen to compare their cluster quality. The results show that the quality of both proposed clustering algorithms of these two factors is close to k-means clustering.