Likelihood Ratio Confidence Sets for Sequential Decision Making
Nicolas Emmenegger, Mojmir Mutny, Andreas Krause
摘要
Certifiable, adaptive uncertainty estimates for unknown quantities are an essential ingredient of sequential decision-making algorithms. Standard approaches rely on problem-dependent concentration results and are limited to a specific combination of parameterization, noise family, and estimator. In this paper, we revisit the likelihood-based inference principle and propose to use likelihood ratios to construct any-time valid confidence sequences without requiring specialized treatment in each application scenario. Our method is especially suitable for problems with well-specified likelihoods, and the resulting sets always maintain the prescribed coverage in a model-agnostic manner. The size of the sets depends on a choice of estimator sequence in the likelihood ratio. We discuss how to provably choose the best sequence of estimators and shed light on connections to online convex optimization with algorithms such as Follow-the-Regularized-Leader. To counteract the initially large bias of the estimators, we propose a reweighting scheme that also opens up deployment in non-parametric settings such as RKHS function classes. We provide a non-asymptotic analysis of the likelihood ratio confidence sets size for generalized linear models, using insights from convex duality and online learning. We showcase the practical strength of our method on generalized linear bandit problems, survival analysis, and bandits with various additive noise distributions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- A Unified Confidence Sequence for Generalized Linear Models, with Applications to BanditsJunghyun Lee, Se-Young Yun, Kwang-Sung JunNeurIPS 2024 · 被引用 35 次
- Transductive Active Learning: Theory and ApplicationsJonas Hübotter, Bhavya Sukhija, Lenart Treven, Yarden As 等NeurIPS 2024 · 被引用 24 次
- Principled Preferential Bayesian OptimizationWenjie Xu, Wenbin Wang, Yuning Jiang, Bratislav Svetozarevic 等ICML 2024 · 被引用 15 次
- Principled Bayesian Optimization in Collaboration with Human ExpertsWenjie Xu, Masaki Adachi, Colin N. Jones, Michael A. OsborneNeurIPS 2024 · 被引用 10 次
- Noise-Adaptive Confidence Sets for Linear Bandits and Application to Bayesian OptimizationKwang-Sung Jun, Jungtaek KimICML 2024 · 被引用 4 次
它引用的顶会 Paper4
- Improved Optimistic Algorithms for Logistic BanditsLouis Faury, Marc Abeille, Clément Calauzènes, Olivier FercoqICML 2020 · 被引用 127 次
- Risk-averse Heteroscedastic Bayesian OptimizationAnastasia Makarova, Ilnura Usmanova, Ilija Bogunovic, Andreas KrauseNeurIPS 2021 · 被引用 47 次
- Stochastic Online Linear Regression: the Forward Algorithm to Replace RidgeReda Ouhamma, Odalric-Ambrym Maillard, Vianney PerchetNeurIPS 2021 · 被引用 18 次
- No-regret Algorithms for Capturing Events in Poisson Point ProcessesMojmir Mutny, Andreas KrauseICML 2021 · 被引用 10 次
相关 Paper
- Improved Algorithms for Stochastic Linear Bandits Using Tail Bounds for Martingale MixturesHamish Flynn, David Reeb, Melih Kandemir, Jan R. PetersNeurIPS 2023 · 被引用 14 次
- Statistical Inference with M-Estimators on Adaptively Collected DataKelly W. Zhang, Lucas Janson, Susan A. MurphyNeurIPS 2021 · 被引用 66 次
- Confidence sequences for sampling without replacementIan Waudby-Smith, Aaditya RamdasNeurIPS 2020 · 被引用 57 次
- Discounted Adaptive Online Learning: Towards Better RegularizationZhiyu Zhang, David Bombara, Heng YangICML 2024 · 被引用 13 次
- Meta-Learning Hypothesis Spaces for Sequential Decision-makingParnian Kassraie, Jonas Rothfuss, Andreas KrauseICML 2022 · 被引用 6 次
