Pith. sign in

REVIEW 1 cited by

Large-scale Multiple Testing: Fundamental Limits of False Discovery Rate Control and Compound Oracle

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2302.06809 v3 pith:DYERHMV2 submitted 2023-02-14 math.ST stat.MEstat.TH

classification math.STstat.MEstat.TH
keywords falseoptimalrulestradeoffdiscoveryproportionrateseparable
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The false discovery rate (FDR) and the false non-discovery rate (FNR), defined as the expected false discovery proportion (FDP) and the false non-discovery proportion (FNP), are the most popular benchmarks for multiple testing. Despite the theoretical and algorithmic advances in recent years, the optimal tradeoff between the FDR and the FNR has been largely unknown except for certain restricted classes of decision rules, e.g., separable rules, or for other performance metrics, e.g., the marginal FDR and the marginal FNR (mFDR and mFNR). In this paper, we determine the asymptotically optimal FDR-FNR tradeoff under the two-group random mixture model when the number of hypotheses tends to infinity. Distinct from the optimal mFDR-mFNR tradeoff, which is achieved by separable decision rules, the optimal FDR-FNR tradeoff requires compound rules even in the large-sample limit and for models as simple as the Gaussian location model. This suboptimality of separable rules also holds for other objectives, such as maximizing the expected number of true discoveries. Finally, to address the limitation of the FDR which only controls the expectation but not the fluctuation of the FDP, we also determine the optimal tradeoff when the FDP is controlled with high probability and show it coincides with that of the mFDR and the mFNR. Extensions to models with a fixed non-null proportion are also obtained.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Besting Good--Turing: Optimality of Non-Parametric Maximum Likelihood for Distribution Estimation

    math.ST 2025-09 conditional novelty 7.0 of 10

    An NPMLE-based empirical Bayes estimator is shown to be competitively optimal (up to log factors) for KL-risk distribution estimation, while Good-Turing is provably suboptimal.

Pith tools