9. Clustering and Nonnegative Matrix Factorization
Lars Eldén · Society for Industrial and Applied Mathematics eBooks · 2007
An important method for data compression and classification is to organize data points in clusters. A cluster is a subset of the set of data points that are close together in some distance measure. One can compute the mean value of each cluster separately and use the means as representatives of the clusters. Equivalently, the means can be used as basis vectors, and all the data points represented by their coordinates with respect to this basis.