PABBO: Preferential Amortized Black-Box Optimization
Xinyu Zhang, Daolang Huang, Samuel Kaski, Julien Martinelli
Abstract
Preferential Bayesian Optimization (PBO) is a sample-efficient method to learn latent user utilities from preferential feedback over a pair of designs. It relies on a statistical surrogate model for the latent function, usually a Gaussian process, and an acquisition strategy to select the next candidate pair to get user feedback on. Due to the non-conjugacy of the associated likelihood, every PBO step requires a significant amount of computations with various approximate inference techniques. This computational overhead is incompatible with the way humans interact with computers, hindering the use of PBO in real-world cases. Building on the recent advances of amortized BO, we propose to circumvent this issue by fully amortizing PBO, meta-learning both the surrogate and the acquisition function. Our method comprises a novel transformer neural process architecture, trained using reinforcement learning and tailored auxiliary losses. On a benchmark composed of synthetic and real-world datasets, our method is several orders of magnitude faster than the usual Gaussian process-based strategies and often outperforms them in accuracy. * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 41191efc-9086-4a6e-b1db-d1060a4bd496Cited by top-tier papers4
- ALINE: Joint Amortization for Bayesian Inference and Active Data AcquisitionDaolang Huang, Xinyi Wen, Ayush Bharti, Samuel Kaski et al.NeurIPS 2025 · 8 citations
- In-Context Multi-Objective OptimizationXinyu Zhang, Conor Hassan, Julien Martinelli, Daolang Huang et al.ICLR 2026 · 6 citations
- JADAI: Jointly Amortizing Adaptive Design and Bayesian InferenceNiels Bracher, Lars Kühmichel, Desi Ivanova, Xavier Intes et al.ICML 2026 · 2 citations
- Bayesian Optimization with Preference Exploration using a Monotonic Neural Network EnsembleHanyang Wang, Juergen Branke, Matthias PoloczekNeurIPS 2025
Builds on24
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- BoTorch: A Framework for Efficient Monte-Carlo Bayesian OptimizationMaximilian Balandat, Brian Karrer, Daniel R. Jiang, Samuel Daulton et al.NeurIPS 2020 · 686 citations
- Transformers Can Do Bayesian InferenceSamuel Müller, Noah Hollmann, Sebastian Pineda-Arango, Josif Grabocka et al.ICLR 2022 · 287 citations
- Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence ModelingTung Nguyen, Aditya GroverICML 2022 · 148 citations
Related papers
- End-to-End Meta-Bayesian Optimisation with Transformer Neural ProcessesAlexandre Maraval, Matthieu Zimmer, Antoine Grosnit, Haitham Bou-AmmarNeurIPS 2023 · 41 citations
- Meta-Learning Acquisition Functions for Transfer Learning in Bayesian OptimizationMichael Volpp, Lukas P. Fröhlich, Kirsten Fischer, Andreas Doerr et al.ICLR 2020 · 104 citations
- Efficient Visual Appearance Optimization by Learning from Prior PreferencesZhipeng Li, Yi-Chi Liao, Christian HolzUIST 2025
- Generative Bayesian Optimization: Generative Models as Acquisition FunctionsRafael Oliveira, Daniel M. Steinberg, Edwin V. BonillaICLR 2026 · 3 citations
- Supporting High-Stakes Decision Making Through Interactive Preference Elicitation in the Latent SpaceMichael Eichelbeck, Tim Voigt, Matthias AlthoffICLR 2026
