On choosing mixture components via non-local priors

Fúquene, Jairo; Steel, Mark; Rossell, David

Full-text links:

Download:

Current browse context:

stat.ME

< prev | next >

new | recent | 1604

Statistics > Methodology

Title: On choosing mixture components via non-local priors

Authors: Jairo Fúquene, Mark Steel, David Rossell

(Submitted on 1 Apr 2016 (v1), last revised 11 Jun 2019 (this version, v5))

Abstract: Choosing the number of mixture components remains an elusive challenge. Model selection criteria can be either overly liberal or conservative and return poorly-separated components of limited practical use. We formalize non-local priors (NLPs) for mixtures and show how they lead to well-separated components with non-negligible weight, interpretable as distinct subpopulations. We also propose an estimator for posterior model probabilities under local and non-local priors, showing that Bayes factors are ratios of posterior to prior empty-cluster probabilities. The estimator is widely applicable and helps set thresholds to drop unoccupied components in overfitted mixtures. We suggest default prior parameters based on multi-modality for Normal/T mixtures and minimal informativeness for categorical outcomes. We characterise theoretically the NLP-induced sparsity, derive tractable expressions and algorithms. We fully develop Normal, Binomial and product Binomial mixtures but the theory, computation and principles hold more generally. We observed a serious lack of sensitivity of the Bayesian information criterion (BIC), insufficient parsimony of the AIC and a local prior, and a mixed behavior of the singular BIC. We also considered overfitted mixtures, their performance was competitive but depended on tuning parameters. Under our default prior elicitation NLPs offered a good compromise between sparsity and power to detect meaningfully-separated components.

Subjects:	Methodology (stat.ME)
Cite as:	arXiv:1604.00314 [stat.ME]
	(or arXiv:1604.00314v5 [stat.ME] for this version)

Submission history

From: Jairo Fúquene [view email]
[v1] Fri, 1 Apr 2016 16:22:00 GMT (187kb,D)
[v2] Wed, 19 Apr 2017 22:00:51 GMT (501kb,D)
[v3] Wed, 13 Jun 2018 14:26:00 GMT (518kb,D)
[v4] Fri, 15 Jun 2018 16:37:09 GMT (518kb,D)
[v5] Tue, 11 Jun 2019 16:47:19 GMT (566kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> stat > arXiv:1604.00314

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Statistics > Methodology

Title: On choosing mixture components via non-local priors

Submission history