Exploring and Exploiting Model Uncertainty in Bayesian Optimization
Zishi Zhang, Tao Ren, Yijie Peng
摘要
In this work, we consider the problem of Bayesian Optimization (BO) under reward model uncertainty —that is, when the underlying distribution type of the reward is unknown and potentially intractable to specify. This challenge is particularly evident in many modern applications, where the reward distribution is highly ill-behaved, often non-stationary, multi-modal, or heavy-tailed. In such settings, classical Gaussian Process (GP)-based BO methods often fail due to their strong modeling assumptions. To address this challenge, we propose a novel surrogate model, the infinity-Gaussian Process ( ∞ -GP), which represents a sequential spatial Dirichlet Process mixture with a GP baseline. The ∞ -GP quantifies both value uncertainty and model uncertainty, enabling more flexible modeling of complex reward structures. Combined with Thompson Sampling, the ∞ -GP facilitates principled exploration and exploitation in the distributional space of reward models. Theoretically, we prove that the ∞ -GP surrogate model can approximate a broad class of reward distributions by effectively exploring the distribution space, achieving near-minimax-optimal posterior contraction rates. Empirically, our method outperforms state-of-the-art approaches in various challenging scenarios, including highly non-stationary and heavy-tailed reward settings where classical GP-based BO often fails.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Unexpected Improvements to Expected Improvement for Bayesian OptimizationSebastian Ament, Samuel Daulton, David Eriksson, Maximilian Balandat 等NeurIPS 2023 · 被引用 280 次
- Efficiently sampling functions from Gaussian process posteriorsJames T. Wilson, Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky 等ICML 2020 · 被引用 186 次
- RLPrompt: Optimizing Discrete Text Prompts with Reinforcement LearningMingkai Deng, Jianyu Wang, Cheng-Ping Hsieh, Yihan Wang 等EMNLP 2022 · 被引用 141 次
- Misspecified Gaussian Process Bandit OptimizationIlija Bogunovic, Andreas KrauseNeurIPS 2021 · 被引用 69 次
相关 Paper
- BayeSQP: Bayesian Optimization through Sequential Quadratic ProgrammingPaul Brunzema, Sebastian TrimpeNeurIPS 2025 · 被引用 7 次
- Modulating Surrogates for Bayesian OptimizationErik Bodin, Markus Kaiser, Ieva Kazlauskaite, Zhenwen Dai 等ICML 2020 · 被引用 11 次
- Policy Search via Bayesian Optimization with Temporal Difference Gaussian ProcessesArmin Lederer, Anuj Srivastava, Marco Bagatella, Andreas KrauseICML 2026
- Thompson Sampling in Function Spaces via Neural OperatorsRafael Oliveira, Xuesong Wang, Kian Ming A. Chai, Edwin V. BonillaNeurIPS 2025 · 被引用 2 次
- Objective Bound Conditional Gaussian Process for Bayesian OptimizationTaewon Jeong, Heeyoung KimICML 2021 · 被引用 3 次
