Bayesian Regularization for Normal Mixture Estimation and Model-Based Clustering

Publication Type:Journal Article
Year of Publication:2007
Authors:C. Fraley, Raftery A. E.
Journal:Journal of Classification
Volume:24
Pagination:155–181
Date Published:September
ISSN:0176-4268
Keywords:classification, cluster, models, taxonomy
Abstract:

Abstract  Normal mixture models are widely used for statistical modeling of data, including cluster analysis. However maximum likelihood estimation (MLE) for normal mixtures using the EM algorithm may fail as the result of singularities or degeneracies. To avoid this, we propose replacing the MLE by a maximum a posteriori (MAP) estimator, also found by the EM algorithm. For choosing the number of components and the model parameterization, we propose a modified version of BIC, where the likelihood is evaluated at the MAP instead of the MLE. We use a highly dispersed proper conjugate prior, containing a small fraction of one observation's worth of information. The resulting method avoids degeneracies and singularities, but when these are not present it gives similar results to the standard method using MLE, EM and BIC.

URL:http://dx.doi.org/10.1007/s00357-007-0004-5
DOI:10.1007/s00357-007-0004-5
Scratchpads developed and conceived by (alphabetical): Ed Baker, Katherine Bouton Alice Heaton Dimitris Koureas, Laurence Livermore, Dave Roberts, Simon Rycroft, Ben Scott, Vince Smith