On the number of components in a Gaussian mixture model
Geoffrey John McLachlan, Suren I. Rathnayake · Wiley Interdisciplinary Reviews Data Mining and Knowledge Discovery · 2014
Mixture distributions, in particular normal mixtures, are applied to data with two main purposes in mind. One is to provide an appealing semiparametric framework in which to model unknown distributional shapes, as an alternative to, say, the kernel density method. The other is to use the mixture model to provide a probabilistic clustering of the data intogclusters corresponding to thegcomponents in the mixture model. In both situations, there is the question of how many components to include in the normal mixture model. We review various methods that have been proposed to answer this question.WIREs Data Mining Knowl Discov2014, 4:341–355. doi: 10.1002/widm.1135 This article is categorized under: Technologies > Machine Learning