Reinforced Few-Shot Acquisition Function Learning for Bayesian Optimization
Bing-Jing Hsieh, Ping-Chun Hsieh, Xi Liu
摘要
Bayesian optimization (BO) conventionally relies on handcrafted acquisition functions (AFs) to sequentially determine the sample points. However, it has been widely observed in practice that the best-performing AF in terms of regret can vary significantly under different types of black-box functions. It has remained a challenge to design one AF that can attain the best performance over a wide variety of black-box functions. This paper aims to attack this challenge through the perspective of reinforced few-shot AF learning (FSAF). Specifically, we first connect the notion of AFs with Q-functions and view a deep Q-network (DQN) as a surrogate differentiable AF. While it serves as a natural idea to combine DQN and an existing few-shot learning method, we identify that such a direct combination does not perform well due to severe overfitting, which is particularly critical in BO due to the need of a versatile sampling policy. To address this, we present a Bayesian variant of DQN with the following three features: (i) It learns a distribution of Q-networks as AFs based on the Kullback-Leibler regularization framework. This inherently provides the uncertainty required in sampling for BO and mitigates overfitting. (ii) For the prior of the Bayesian DQN, we propose to use a demo policy induced by an off-the-shelf AF for better training stability. (iii) On the meta-level, we leverage the meta-loss of Bayesian model-agnostic meta-learning, which serves as a natural companion to the proposed FSAF. Moreover, with the proper design of the Q-networks, FSAF is general-purpose in that it is agnostic to the dimension and the cardinality of the input domain. Through extensive experiments, we demonstrate that the FSAF achieves comparable or better regrets than the state-of-the-art benchmarks on a wide variety of synthetic and real-world test functions. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- End-to-End Meta-Bayesian Optimisation with Transformer Neural ProcessesAlexandre Maraval, Matthieu Zimmer, Antoine Grosnit, Haitham Bou-AmmarNeurIPS 2023 · 被引用 41 次
- Continual Human-in-the-Loop OptimizationYi-Chi Liao, Paul Streli, Zhipeng Li, Christoph Gebhardt 等CHI 2025 · 被引用 11 次
- A Meta-Bayesian Approach for Rapid Online Parametric Optimization for Wrist-based InteractionsYi-Chi Liao, Ruta Desai, Alec M. Pierce, Krista E. Taylor 等CHI 2024 · 被引用 6 次
- Adaptive Acquisition Selection for Bayesian Optimization with Large Language ModelsGiang Ngo, Dat Phan Trong, Dang Nguyen, Sunil Gupta 等ICLR 2026 · 被引用 6 次
- Efficient Human-in-the-Loop Optimization via Priors Learned from User ModelsYi-Chi Liao, João Marcelo Evangelista Belo, Hee-Seung Moon, Jürgen Steimle 等CHI 2026 · 被引用 3 次
它引用的顶会 Paper5
- Bayesian Meta-Learning for the Few-Shot Setting via Deep KernelsMassimiliano Patacchiola, Jack Turner, Elliot J. Crowley, Michael F. P. O'Boyle 等NeurIPS 2020 · 被引用 167 次
- Meta-Learning Acquisition Functions for Transfer Learning in Bayesian OptimizationMichael Volpp, Lukas P. Fröhlich, Kirsten Fischer, Andreas Doerr 等ICLR 2020 · 被引用 104 次
- Few-Shot Bayesian Optimization with Deep Kernel SurrogatesMartin Wistuba, Josif GrabockaICLR 2021 · 被引用 87 次
- BINOCULARS for efficient, nonmyopic sequential experimental designShali Jiang, Henry Chai, Javier González, Roman GarnettICML 2020 · 被引用 56 次
- Efficient Nonmyopic Bayesian Optimization via One-Shot Multi-Step TreesShali Jiang, Daniel R. Jiang, Maximilian Balandat, Brian Karrer 等NeurIPS 2020 · 被引用 54 次
相关 Paper
- FSEO: Few-Shot Evolutionary Optimization via Meta-Learning for Expensive Multi-Objective OptimizationXunzhao YuNeurIPS 2025
- MALIBO: Meta-learning for Likelihood-free Bayesian OptimizationJiarong Pan, Stefan Falkner, Felix Berkenkamp, Joaquin VanschorenICML 2024 · 被引用 2 次
- Bayesian Optimization over Discrete and Mixed Spaces via Probabilistic ReparameterizationSamuel Daulton, Xingchen Wan, David Eriksson, Maximilian Balandat 等NeurIPS 2022 · 被引用 71 次
- Shallow Bayesian Meta Learning for Real-World Few-Shot RecognitionXueting Zhang, Debin Meng, Henry Gouk, Timothy M. HospedalesICCV 2021 · 被引用 88 次
- BOFormer: Learning to Solve Multi-Objective Bayesian Optimization via Non-Markovian RLYu-Heng Hung, Kai-Jie Lin, Yu-Heng Lin, Chien-Yi Wang 等ICLR 2025
