Comparative Analysis of Hybrid Clustering Algorithm on Different Dataset

Hira Malik, Noor-u-Zaman Laghari, Din Muhammad Sangrasi, Zulfiqar Ali Dayo · 2018

Clustering is a data mining technique, in which data is grouped based on similarity and dissimilarity. Clustering is usually used to identify hidden pattern in multidimensional complex data and, these hidden pattern provide bases for making decisions. The objective of this research to find the best clustering algorithm. K-Mean is a famous clustering algorithm, which is simple and easy to implement, but the drawback of K-Mean is that, it does not work with higher dimensional data, for aiding with this drawback K-Mean is fused with other clustering algorithms such as PSO (Particle Swarm optimization) and PCA (Principle Component Analysis) for better results and cluster identifications. In this paper authors used hybrid clustering approach (K-Mean, PSO-K-Mean and PCA-K-Mean) to improve the clustering result based on parameter (Purity, Rand index and Computation Time) on different data sets taken from UCI Repository.

Read the paper · More papers on PaperTik