An Expansion of X-Means for Automatically Determining the Optimal Number of Clusters a^EUR" Progressive Iterations of K-Means and Merging of the Clusters.
Tsunenori Ishioka · Computational intelligence · 2005
We expand a non-hierarchical clustering algorithm that can determine the optimal number of clusters by using iterations of -means and a stopping rule based on Bayesian Information Criterion (BIC). The procedure requires merging the clusters that a -means iteration has made to avoid unsuitable division caused by the division order. By using this additional merging operation, the case of adequate clustering was increased for various types of simulation runs. With no prior information about the number of clusters, our method can get the optimal clustering based on information theory instead of on a heuristic method. The computational complexity of our method is for the sample size and the number of final clusters, .