Learning to Bet for Horizon-Aware Anytime-Valid Testing
Ege Onur Taga, Samet Oymak, Shubhanshu Shekhar
摘要
We develop horizon-aware anytime-valid tests and confidence sequences for bounded means under a strict deadline . Using the betting/e-process framework, we cast horizon-aware betting as a finite-horizon optimal control problem with state space , where is the time and is the test martingale value. We first show that in certain interior regions of the state space, policies that deviate significantly from Kelly betting are provably suboptimal, while Kelly betting reaches the threshold with high probability. We then identify sufficient conditions showing that outside this region, more aggressive betting than Kelly can be better if the bettor is behind schedule, and less aggressive can be better if the bettor is ahead. Taken together these results suggest a simple phase diagram in the plane, delineating regions where Kelly, fractional Kelly, and aggressive betting may be preferable. Guided by this phase diagram, we introduce a Deep Reinforcement Learning approach based on a universal Deep Q-Network (DQN) agent that learns a single policy from synthetic experience and maps simple statistics of past observations to bets across horizons and null values. In limited-horizon experiments, the learned DQN policy yields state-of-the-art results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Anytime Detection of Strategic Deviations in Multi-Agent SystemsEtienne Gauthier, Francis Bach, Michael JordanICML 2026 · 被引用 2 次
- Variational Bayesian Reinforcement Learning with Regret BoundsBrendan O'DonoghueNeurIPS 2021 · 被引用 48 次
- Sequential Predictive Two-Sample and Independence TestingAleksandr Podkopaev, Aaditya RamdasNeurIPS 2023 · 被引用 29 次
- Learning to Stop: Deep Learning for Mean Field Optimal StoppingLorenzo Magnino, Yuchen Zhu, Mathieu LaurièreICML 2025
- Protected Test-Time Adaptation via Online Entropy Matching: A Betting ApproachYarin Bar, Shalev Shaer, Yaniv RomanoNeurIPS 2024 · 被引用 27 次
