End-to-End Meta-Bayesian Optimisation with Transformer Neural Processes
Alexandre Maraval, Matthieu Zimmer, Antoine Grosnit, Haitham Bou-Ammar
摘要
Meta-Bayesian optimisation (meta-BO) aims to improve the sample efficiency of Bayesian optimisation by leveraging data from related tasks. While previous methods successfully meta-learn either a surrogate model or an acquisition function independently, joint training of both components remains an open challenge. This paper proposes the first end-to-end differentiable meta-BO framework that generalises neural processes to learn acquisition functions via transformer architectures. We enable this end-to-end framework with reinforcement learning (RL) to tackle the lack of labelled acquisition data. Early on, we notice that training transformer-based neural processes from scratch with RL is challenging due to insufficient supervision, especially when rewards are sparse. We formalise this claim with a combinatorial analysis showing that the widely used notion of regret as a reward signal exhibits a logarithmic sparsity pattern in trajectory lengths. To tackle this problem, we augment the RL objective with an auxiliary task that guides part of the architecture to learn a valid probabilistic model as an inductive bias. We demonstrate that our method achieves state-of-the-art regret results against various baselines in experiments on standard hyperparameter optimisation tasks and also outperforms others in the real-world problems of mixed-integer programming tuning, antibody design, and logic synthesis for electronic design automation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- All-in-one simulation-based inferenceManuel Glöckler, Michael Deistler, Christian Dietrich Weilbach, Frank Wood 等ICML 2024 · 被引用 74 次
- Amortized Bayesian Experimental Design for Decision-MakingDaolang Huang, Yujia Guo, Luigi Acerbi, Samuel KaskiNeurIPS 2024 · 被引用 24 次
- ALINE: Joint Amortization for Bayesian Inference and Active Data AcquisitionDaolang Huang, Xinyi Wen, Ayush Bharti, Samuel Kaski 等NeurIPS 2025 · 被引用 8 次
- In-Context Multi-Objective OptimizationXinyu Zhang, Conor Hassan, Julien Martinelli, Daolang Huang 等ICLR 2026 · 被引用 6 次
- MALIBO: Meta-learning for Likelihood-free Bayesian OptimizationJiarong Pan, Stefan Falkner, Felix Berkenkamp, Joaquin VanschorenICML 2024 · 被引用 2 次
它引用的顶会 Paper12
- Stabilizing Transformers for Reinforcement LearningEmilio Parisotto, H. Francis Song, Jack W. Rae, Razvan Pascanu 等ICML 2020 · 被引用 464 次
- Transformers Can Do Bayesian InferenceSamuel Müller, Noah Hollmann, Sebastian Pineda-Arango, Josif Grabocka 等ICLR 2022 · 被引用 287 次
- Convolutional Conditional Neural ProcessesJonathan Gordon, Wessel P. Bruinsma, Andrew Y. K. Foong, James Requeima 等ICLR 2020 · 被引用 200 次
- Bayesian Meta-Learning for the Few-Shot Setting via Deep KernelsMassimiliano Patacchiola, Jack Turner, Elliot J. Crowley, Michael F. P. O'Boyle 等NeurIPS 2020 · 被引用 167 次
- Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence ModelingTung Nguyen, Aditya GroverICML 2022 · 被引用 148 次
相关 Paper
- PABBO: Preferential Amortized Black-Box OptimizationXinyu Zhang, Daolang Huang, Samuel Kaski, Julien MartinelliICLR 2025
- Meta-Learning Acquisition Functions for Transfer Learning in Bayesian OptimizationMichael Volpp, Lukas P. Fröhlich, Kirsten Fischer, Andreas Doerr 等ICLR 2020 · 被引用 104 次
- ProxyBO: Accelerating Neural Architecture Search via Bayesian Optimization with Zero-Cost ProxiesYu Shen, Yang Li, Jian Zheng, Wentao Zhang 等AAAI 2023 · 被引用 43 次
- Reinforced Few-Shot Acquisition Function Learning for Bayesian OptimizationBing-Jing Hsieh, Ping-Chun Hsieh, Xi LiuNeurIPS 2021 · 被引用 29 次
- Estimating Interventional Distributions with Uncertain Causal Graphs through Meta-LearningAnish Dhir, Cristiana Diaconu, Valentinian Lungu, James Requeima 等NeurIPS 2025 · 被引用 16 次
