-PFN: Fast Entropy Search via In-Context Learning
Herilalaina Rakotoarison, Steven Adriaensen, Tom Viering, Carl Hvarfner, Samuel Gabriel Müller, Frank Hutter, Eytan Bakshy
Abstract
Information-theoretic acquisition functions such as Entropy Search (ES) offer a principled exploration-exploitation framework for Bayesian optimization (BO). However, their practical implementation relies on complicated and slow approximations, i.e., a Monte Carlo estimation of the information gain. This complexity can introduce numerical errors and requires specialized, hand-crafted implementations. We propose a two-stage amortization strategy that learns to approximate entropy search-based acquisition functions using Prior-data Fitted Networks (PFNs) in a single forward pass. A first PFN is trained to be conditioned on information about the optima; second, the α-PFN is trained to predict the expected information gain by training on information gains measured with the first PFN. The α-PFN offers a flexible learned approximation, which replaces the complex heuristic approximations with a single forward pass per candidate, enabling rapid and extensible acquisition evaluation. Empirically, our approach is competitive with state-of-the-art entropy search implementations on synthetic and real-world benchmarks, while accelerating the different entropy search variants across all our experiments, with speed ups over 50x. Source code: https://github. com/automl/AlphaPFN .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 040229af-a84b-4bd8-babc-1c824b14ad7cBuilds on23
- BoTorch: A Framework for Efficient Monte-Carlo Bayesian OptimizationMaximilian Balandat, Brian Karrer, Daniel R. Jiang, Samuel Daulton et al.NeurIPS 2020 · 686 citations
- Transformers Can Do Bayesian InferenceSamuel Müller, Noah Hollmann, Sebastian Pineda-Arango, Josif Grabocka et al.ICLR 2022 · 287 citations
- ForecastPFN: Synthetically-Trained Zero-Shot ForecastingSamuel Dooley, Gurnoor Singh Khurana, Chirag Mohapatra, Siddartha V. Naidu et al.NeurIPS 2023 · 142 citations
- Local Latent Space Bayesian Optimization over Structured InputsNatalie Maus, Haydn Thomas Jones, Juston Moore, Matt J. Kusner et al.NeurIPS 2022 · 118 citations
- Towards Learning Universal Hyperparameter Optimizers with TransformersYutian Chen, Xingyou Song, Chansoo Lee, Zi Wang et al.NeurIPS 2022 · 106 citations
Related papers
- A Unified Framework for Entropy Search and Expected Improvement in Bayesian OptimizationNuojin Cheng, Leonard Papenmeier, Stephen Becker, Luigi NardiICML 2025
- PFNs4BO: In-Context Learning for Bayesian OptimizationSamuel Müller, Matthias Feurer, Noah Hollmann, Frank HutterICML 2023 · 71 citations
- Bayesian Optimization of Function Networks with Partial EvaluationsPoompol Buathong, Jiayue Wan, Raul Astudillo, Samuel Daulton et al.ICML 2024 · 10 citations
- Joint Entropy Search For Maximally-Informed Bayesian OptimizationCarl Hvarfner, Frank Hutter, Luigi NardiNeurIPS 2022 · 69 citations
- Generalizing Bayesian Optimization with Decision-theoretic EntropiesWillie Neiswanger, Lantao Yu, Shengjia Zhao, Chenlin Meng et al.NeurIPS 2022 · 15 citations
