References & Citations
Statistics > Methodology
Title: On choosing mixture components via non-local priors
(Submitted on 1 Apr 2016 (v1), last revised 11 Jun 2019 (this version, v5))
Abstract: Choosing the number of mixture components remains an elusive challenge. Model selection criteria can be either overly liberal or conservative and return poorly-separated components of limited practical use. We formalize non-local priors (NLPs) for mixtures and show how they lead to well-separated components with non-negligible weight, interpretable as distinct subpopulations. We also propose an estimator for posterior model probabilities under local and non-local priors, showing that Bayes factors are ratios of posterior to prior empty-cluster probabilities. The estimator is widely applicable and helps set thresholds to drop unoccupied components in overfitted mixtures. We suggest default prior parameters based on multi-modality for Normal/T mixtures and minimal informativeness for categorical outcomes. We characterise theoretically the NLP-induced sparsity, derive tractable expressions and algorithms. We fully develop Normal, Binomial and product Binomial mixtures but the theory, computation and principles hold more generally. We observed a serious lack of sensitivity of the Bayesian information criterion (BIC), insufficient parsimony of the AIC and a local prior, and a mixed behavior of the singular BIC. We also considered overfitted mixtures, their performance was competitive but depended on tuning parameters. Under our default prior elicitation NLPs offered a good compromise between sparsity and power to detect meaningfully-separated components.
Submission history
From: Jairo Fúquene [view email][v1] Fri, 1 Apr 2016 16:22:00 GMT (187kb,D)
[v2] Wed, 19 Apr 2017 22:00:51 GMT (501kb,D)
[v3] Wed, 13 Jun 2018 14:26:00 GMT (518kb,D)
[v4] Fri, 15 Jun 2018 16:37:09 GMT (518kb,D)
[v5] Tue, 11 Jun 2019 16:47:19 GMT (566kb,D)
Link back to: arXiv, form interface, contact.