Large-scale multiple testing: Fundamental limits of false discovery rate control and compound oracle

Yutong Nie & Yihong Wu

Annals of Statistics2026https://doi.org/10.1214/25-aos2581article
AJG 4*ABDC A*
Weight
0.50

What the paper says

The false discovery rate (FDR) and the false nondiscovery rate (FNR), defined as the expected false discovery proportion (FDP) and the false nondiscovery proportion (FNP), are the most popular benchmarks for multiple testing. Despite the theoretical and algorithmic advances in recent years, the optimal trade-off between the FDR and the FNR has been largely unknown, except for certain restricted classes of decision rules, for example, separable rules, or for other performance metrics, for example, the marginal FDR and the marginal FNR (mFDR and mFNR). In this paper we determine the asymptotically optimal FDR-FNR trade-off under the two-group random mixture model when the number of hypotheses tends to infinity. Distinct from the optimal mFDR-mFNR trade-off, which is achieved by separable decision rules, the optimal FDR-FNR trade-off requires compound rules, even in the large-sample limit and for models as simple as the Gaussian location model. This suboptimality of separable rules also holds for other objectives, such as maximizing the expected number of true discoveries. Finally, to address the limitation of the FDR, which only controls the expectation but not the fluctuation of the FDP, we also determine the optimal tradeoff when the FDP is controlled with high probability and show it coincides with that of the mFDR and the mFNR. Extensions to models with a fixed nonnull proportion are also obtained.

Open paper page →

Cite this paper

https://doi.org/https://doi.org/10.1214/25-aos2581

Or copy a formatted citation

@article{yutong2026,
  title        = {{Large-scale multiple testing: Fundamental limits of false discovery rate control and compound oracle}},
  author       = {Yutong Nie & Yihong Wu},
  journal      = {Annals of Statistics},
  year         = {2026},
  doi          = {https://doi.org/https://doi.org/10.1214/25-aos2581},
}

Paste directly into BibTeX, Zotero, or your reference manager.

Flag this paper

Large-scale multiple testing: Fundamental limits of false discovery rate control and compound oracle

Flags are reviewed by the Arbiter methodology team within 5 business days.


Evidence weight

0.50

Balanced mode · F 0.40 / M 0.15 / V 0.05 / R 0.40

F · citation impact0.50 × 0.4 = 0.20
M · momentum0.50 × 0.15 = 0.07
V · venue signal0.50 × 0.05 = 0.03
R · text relevance †0.50 × 0.4 = 0.20

† Text relevance is estimated at 0.50 on the detail page — for your query’s actual relevance score, open this paper from a search result.