A unified framework for bandit multiple testing
Ziyu Xu, Ruodu Wang, Aaditya Ramdas
摘要
In bandit multiple hypothesis testing, each arm corresponds to a different null hypothesis that we wish to test, and the goal is to design adaptive algorithms that correctly identify large set of interesting arms (true discoveries), while only mistakenly identifying a few uninteresting ones (false discoveries). One common metric in non-bandit multiple testing is the false discovery rate (FDR). We propose a unified, modular framework for bandit FDR control that emphasizes the decoupling of exploration and summarization of evidence. We utilize the powerful martingale-based concept of "e-processes" to ensure FDR control for arbitrary composite nulls, exploration rules and stopping times in generic problem settings. In particular, valid FDR control holds even if the reward distributions of the arms could be dependent, multiple arms may be queried simultaneously, and multiple (cooperating or competing) agents may be querying arms, covering combinatorial semi-bandit type settings as well. Prior work has considered in great detail the setting where each arm's reward distribution is independent and sub-Gaussian, and a single arm is queried at each step. Our framework recovers matching sample complexity guarantees in this special case, and performs comparably or better in practice. For other settings, sample complexities will depend on the finer details of the problem (composite nulls being tested, exploration algorithm, data dependence structure, stopping rule) and we do not explore these; our contribution is to show that the FDR guarantee is clean and entirely agnostic to these details.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Adaptive Identification of Populations with Treatment Benefit in Clinical Trials: Machine Learning Challenges and SolutionsAlicia Curth, Alihan Hüyük, Mihaela van der SchaarICML 2023 · 被引用 3 次
- A New Theoretical Framework for Fast and Accurate Online Decision-MakingNicolò Cesa-Bianchi, Tommaso Cesari, Yishay Mansour, Vianney PerchetNeurIPS 2021 · 被引用 2 次
- Foundations of Testing for Finite-Sample Causal DiscoveryTom Yan, Ziyu Xu, Zachary Chase LiptonICML 2024 · 被引用 1 次
- Adaptive Learn-then-Test: Statistically Valid and Efficient Hyperparameter SelectionMatteo Zecchin, Sangwoo Park, Osvaldo SimeoneICML 2025
- On the Robustness of Bandit Multiple TestingZhengyu Zhou, Weiwei LiuAAAI 2026
它引用的顶会 Paper1
相关 Paper
- PAPRIKA: Private Online False Discovery Rate ControlWanrong Zhang, Gautam Kamath, Rachel CummingsICML 2021 · 被引用 6 次
- AMDP: An Adaptive Detection Procedure for False Discovery Rate Control in High-Dimensional Mediation AnalysisJiarong Ding, Xuehu ZhuNeurIPS 2023 · 被引用 2 次
- Peeking with PEAK: Sequential, Nonparametric Composite Hypothesis Tests for Means of Multiple Data StreamsBrian Cho, Kyra Gan, Nathan KallusICML 2024 · 被引用 14 次
- On the Adversarial Robustness of Benjamini HochbergLouis L. Chen, Roberto Szechtman, Matan SeriNeurIPS 2024 · 被引用 2 次
- Familywise Error Rate Control by Interactive UnmaskingBoyan Duan, Aaditya Ramdas, Larry A. WassermanICML 2020 · 被引用 10 次
