Ensemble sampling for linear bandits: small ensembles suffice
David Janz, Alexander E. Litvak, Csaba Szepesvári
Abstract
We provide the first useful and rigorous analysis of ensemble sampling for the stochastic linear bandit setting. In particular, we show that, under standard assumptions, for a -dimensional stochastic linear bandit with an interaction horizon , ensemble sampling with an ensemble of size of order incurs regret at most of the order . Ours is the first result in any structured setting not to require the size of the ensemble to scale linearly with -- which defeats the purpose of ensemble sampling -- while obtaining near order regret. Our result is also the first to allow for infinite action sets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1c9a1abd-9371-441e-b0dd-3a1cb676a085Cited by top-tier papers2
- Improved Regret of Linear Ensemble SamplingHarin Lee, Min-hwan OhNeurIPS 2024 · 8 citations
- Scalable Exploration via Ensemble++Yingru Li, Jiawei Xu, Baoxiang Wang, Zhi-Quan Tom LuoNeurIPS 2025
Builds on5
- Efficient Model-Based Reinforcement Learning through Optimistic Policy Search and PlanningSebastian Curi, Felix Berkenkamp, Andreas KrauseNeurIPS 2020 · 120 citations
- Hypermodels for ExplorationVikranth Dwaracherla, Xiuyuan Lu, Morteza Ibrahimi, Ian Osband et al.ICLR 2020 · 49 citations
- An Analysis of Ensemble SamplingChao Qin, Zheng Wen, Xiuyuan Lu, Benjamin Van RoyNeurIPS 2022 · 30 citations
- Improved Regret of Linear Ensemble SamplingHarin Lee, Min-hwan OhNeurIPS 2024 · 8 citations
- Batch Ensemble for Variance Dependent Regret in Stochastic BanditsAsaf B. Cassel, Orin Levy, Yishay MansourAAAI 2025 · 3 citations
Related papers
- Stochastic Linear Bandits with Parameter NoiseDaniel Ezer, Alon Peled-Cohen, Yishay MansourICML 2026
- Impact of Representation Learning in Linear BanditsJiaqi Yang, Wei Hu, Jason D. Lee, Simon Shaolei DuICLR 2021 · 58 citations
- Linear Bandits with Feature FeedbackUrvashi Oswal, Aniruddha Bhargava, Robert NowakAAAI 2020 · 6 citations
- On the Suboptimality of Thompson Sampling in High DimensionsRaymond Zhang, Richard CombesNeurIPS 2021 · 6 citations
- Sparsity-Agnostic Linear Bandits with Adaptive AdversariesTianyuan Jin, Kyoungseok Jang, Nicolò Cesa-BianchiNeurIPS 2024 · 2 citations
