Online Prediction of Stochastic Sequences with High Probability Regret Bounds
Matthias Frey, Jonathan H. Manton, Jingge Zhu
Abstract
We revisit the classical problem of universal prediction of stochastic sequences with a finite time horizon known to the learner. The question we investigate is whether it is possible to derive vanishing regret bounds that hold with high probability, complementing existing bounds from the literature that hold in expectation. We propose such high-probability bounds which have a very similar form as the prior expectation bounds. For the case of universal prediction of a stochastic process over a countable alphabet, our bound states a convergence rate of with probability as least compared to prior known in-expectation bounds of the order . We also propose an impossibility result which proves that it is not possible to improve the exponent of in a bound of the same form without making additional assumptions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 71ab6bf0-33ed-461a-8334-b6b12ed70d2cBuilds on2
Related papers
- Online Learning in the Random-Order ModelMartino Bernasconi, Andrea Celli, Riccardo Colini-Baldeschi, Federico Fusco et al.ICML 2025
- Preselection BanditsViktor Bengs, Eyke HüllermeierICML 2020 · 7 citations
- Online Linear Regression in Dynamic Environments via DiscountingAndrew Jacobsen, Ashok CutkoskyICML 2024 · 15 citations
- When Demands Evolve Larger and Noisier: Learning and Earning in a Growing EnvironmentFeng Zhu, Zeyu ZhengICML 2020 · 15 citations
- Improved Regret Bounds for Tracking Experts with MemoryJames Robinson, Mark HerbsterNeurIPS 2021 · 4 citations
