Bias-Robust Bayesian Optimization via Dueling Bandits
Johannes Kirschner, Andreas Krause
Abstract
We consider Bayesian optimization in settings where observations can be adversarially biased, for example by an uncontrolled hidden confounder. Our first contribution is a reduction of the confounded setting to the dueling bandit model. Then we propose a novel approach for dueling bandits based on information-directed sampling (IDS). Thereby, we obtain the first efficient kernelized algorithm for dueling bandits that comes with cumulative regret guarantees. Our analysis further generalizes a previously proposed semi-parametric linear bandit model to non-linear reward functions, and uncovers interesting links to doubly-robust estimation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e1b4556e-5e8a-4363-a8a4-fd363adb6f6bCited by top-tier papers9
- Data-Driven Offline Decision-Making via Invariant Representation LearningHan Qi, Yi Su, Aviral Kumar, Sergey LevineNeurIPS 2022 · 43 citations
- Optimal Design for Human Preference ElicitationSubhojyoti Mukherjee, Anusha Lalitha, Kousha Kalantari, Aniket Deshmukh et al.NeurIPS 2024 · 20 citations
- Robust and Conjugate Gaussian Process RegressionMatías Altamirano, François-Xavier Briol, Jeremias KnoblauchICML 2024 · 18 citations
- A Robust Phased Elimination Algorithm for Corruption-Tolerant Gaussian Process BanditsIlija Bogunovic, Zihan Li, Andreas Krause, Jonathan ScarlettNeurIPS 2022 · 13 citations
- Bandits with Preference Feedback: A Stackelberg Game PerspectiveBarna Pásztor, Parnian Kassraie, Andreas KrauseNeurIPS 2024 · 12 citations
Related papers
- Information Directed Sampling for Sparse Linear BanditsBotao Hao, Tor Lattimore, Wei DengNeurIPS 2021 · 22 citations
- Sparse Optimistic Information Directed SamplingLudovic Schwartz, Hamish Flynn, Gergely NeuNeurIPS 2025 · 1 citation
- Delayed Feedback in Kernel BanditsSattar Vakili, Danyal Ahmed, Alberto Bernacchia, Ciara Pike-BurkeICML 2023 · 8 citations
- Experimental Design for Optimization of Orthogonal Projection Pursuit ModelsMojmir Mutny, Johannes Kirschner, Andreas KrauseAAAI 2020 · 8 citations
- Stochastic Bayesian Optimization with Unknown Continuous Context Distribution via Kernel Density EstimationXiaobin Huang, Lei Song, Ke Xue, Chao QianAAAI 2024 · 3 citations
