Task-Agnostic Amortized Inference of Gaussian Process Hyperparameters
Sulin Liu, Xingyuan Sun, Peter J. Ramadge, Ryan P. Adams
Abstract
Gaussian processes (GPs) are flexible priors for modeling functions. However, their success depends on the kernel accurately reflecting the properties of the data. One of the appeals of the GP framework is that the marginal likelihood of the kernel hyperparameters is often available in closed form, enabling optimization and sampling procedures to fit these hyperparameters to data. Unfortunately, point-wise evaluation of the marginal likelihood is expensive due to the need to solve a linear system; searching or sampling the space of hyperparameters thus often dominates the practical cost of using GPs. We introduce an approach to the identification of kernel hyperparameters in GP regression and related problems that sidesteps the need for costly marginal likelihoods. Our strategy is to "amortize" inference over hyperparameters by training a single neural network, which consumes a set of regression data and produces an estimate of the kernel function, useful across different tasks. To accommodate the varying dimension and cardinality of different regression problems, we use a hierarchical self-attention-based neural network that produces estimates of the hyperparameters which are invariant to the order of the input data points and data dimensions. We show that a single neural model trained on synthetic data is able to generalize directly to several different unseen real-world GP use cases. Our experiments demonstrate that the estimated hyperparameters are comparable in quality to those from the conventional model selection procedures, while being much faster to obtain, significantly accelerating GP regression and its related applications such as Bayesian optimization and Bayesian quadrature. The code and pre-trained model are available at https: //github.com/PrincetonLIPS/AHGP . Gaussian Processes In this section, we establish background concepts and notations necessary for the discussion of the amortized hyperparameters inference approach described in Section 3.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1f39fa1b-2a33-4684-af6e-e9e1b3807977Cited by top-tier papers6
- Neural Diffusion ProcessesVincent Dutordoir, Alan Saul, Zoubin Ghahramani, Fergus SimpsonICML 2023 · 52 citations
- Kernel Identification Through TransformersFergus Simpson, Ian Davies, Vidhi Lalchand, Alessandro Vullo et al.NeurIPS 2021 · 17 citations
- Practical Equivariances via Relational Conditional Neural ProcessesDaolang Huang, Manuel Haussmann, Ulpu Remes, S. T. John et al.NeurIPS 2023 · 14 citations
- AME: Attention and Memory Enhancement in Hyper-Parameter OptimizationNuo Xu, Jianlong Chang, Xing Nie, Chunlei Huo et al.CVPR 2022 · 6 citations
- Revisiting Active Sets for Gaussian Process DecodersPablo Moreno-Muñoz, Cilie W. Feldager, Søren HaubergNeurIPS 2022 · 5 citations
Related papers
- Input Dependent Sparse Gaussian ProcessesBahram Jafrasteh, Carlos Villacampa-Calvo, Daniel Hernández-LobatoICML 2022 · 7 citations
- Marginalised Gaussian Processes with Nested SamplingFergus Simpson, Vidhi Lalchand, Carl Edward RasmussenNeurIPS 2021 · 12 citations
- Empirical Gaussian ProcessesJihao Andreas Lin, Sebastian Ament, Louis Tiao, David Eriksson et al.ICML 2026
- Modeling All Response Surfaces in One for Conditional Search SpacesJiaxing Li, Wei Liu, Chao Xue, Yibing Zhan et al.AAAI 2025 · 1 citation
- Kernel Functional OptimisationArun Kumar Anjanapura Venkatesh, Alistair Shilton, Santu Rana, Sunil Gupta et al.NeurIPS 2021 · 6 citations
