Pith. sign in

REVIEW 1 cited by

Inconsistency of Pitman-Yor process mixtures for the number of components

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1309.0024 v1 pith:BYVXTAW5 submitted 2013-08-30 math.ST stat.MLstat.TH

Inconsistency of Pitman-Yor process mixtures for the number of components

classification math.ST stat.MLstat.TH
keywords componentsnumbermixturesfamiliesprocessdatadpmsfinite
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

In many applications, a finite mixture is a natural model, but it can be difficult to choose an appropriate number of components. To circumvent this choice, investigators are increasingly turning to Dirichlet process mixtures (DPMs), and Pitman-Yor process mixtures (PYMs), more generally. While these models may be well-suited for Bayesian density estimation, many investigators are using them for inferences about the number of components, by considering the posterior on the number of components represented in the observed data. We show that this posterior is not consistent --- that is, on data from a finite mixture, it does not concentrate at the true number of components. This result applies to a large class of nonparametric mixtures, including DPMs and PYMs, over a wide variety of families of component distributions, including essentially all discrete families, as well as continuous exponential families satisfying mild regularity conditions (such as multivariate Gaussians).

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. When (not) to trust Monte Carlo approximations for hierarchical Bayesian inference

    astro-ph.HE 2025-09 conditional novelty 7.0

    A unified error statistic E-hat measures information lost to Monte Carlo noise in hierarchical Bayesian inference, with a recommended cutoff of 0.2 bits.