Task-Agnostic Amortized Inference of Gaussian Process Hyperparameters
Sulin Liu, Xingyuan Sun, Peter J. Ramadge, Ryan P. Adams
摘要
Gaussian processes (GPs) are flexible priors for modeling functions. However, their success depends on the kernel accurately reflecting the properties of the data. One of the appeals of the GP framework is that the marginal likelihood of the kernel hyperparameters is often available in closed form, enabling optimization and sampling procedures to fit these hyperparameters to data. Unfortunately, point-wise evaluation of the marginal likelihood is expensive due to the need to solve a linear system; searching or sampling the space of hyperparameters thus often dominates the practical cost of using GPs. We introduce an approach to the identification of kernel hyperparameters in GP regression and related problems that sidesteps the need for costly marginal likelihoods. Our strategy is to "amortize" inference over hyperparameters by training a single neural network, which consumes a set of regression data and produces an estimate of the kernel function, useful across different tasks. To accommodate the varying dimension and cardinality of different regression problems, we use a hierarchical self-attention-based neural network that produces estimates of the hyperparameters which are invariant to the order of the input data points and data dimensions. We show that a single neural model trained on synthetic data is able to generalize directly to several different unseen real-world GP use cases. Our experiments demonstrate that the estimated hyperparameters are comparable in quality to those from the conventional model selection procedures, while being much faster to obtain, significantly accelerating GP regression and its related applications such as Bayesian optimization and Bayesian quadrature. The code and pre-trained model are available at https: //github.com/PrincetonLIPS/AHGP . Gaussian Processes In this section, we establish background concepts and notations necessary for the discussion of the amortized hyperparameters inference approach described in Section 3.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Neural Diffusion ProcessesVincent Dutordoir, Alan Saul, Zoubin Ghahramani, Fergus SimpsonICML 2023 · 被引用 52 次
- Kernel Identification Through TransformersFergus Simpson, Ian Davies, Vidhi Lalchand, Alessandro Vullo 等NeurIPS 2021 · 被引用 17 次
- Practical Equivariances via Relational Conditional Neural ProcessesDaolang Huang, Manuel Haussmann, Ulpu Remes, S. T. John 等NeurIPS 2023 · 被引用 14 次
- AME: Attention and Memory Enhancement in Hyper-Parameter OptimizationNuo Xu, Jianlong Chang, Xing Nie, Chunlei Huo 等CVPR 2022 · 被引用 6 次
- Revisiting Active Sets for Gaussian Process DecodersPablo Moreno-Muñoz, Cilie W. Feldager, Søren HaubergNeurIPS 2022 · 被引用 5 次
相关 Paper
- Input Dependent Sparse Gaussian ProcessesBahram Jafrasteh, Carlos Villacampa-Calvo, Daniel Hernández-LobatoICML 2022 · 被引用 7 次
- Marginalised Gaussian Processes with Nested SamplingFergus Simpson, Vidhi Lalchand, Carl Edward RasmussenNeurIPS 2021 · 被引用 12 次
- Empirical Gaussian ProcessesJihao Andreas Lin, Sebastian Ament, Louis Tiao, David Eriksson 等ICML 2026
- Modeling All Response Surfaces in One for Conditional Search SpacesJiaxing Li, Wei Liu, Chao Xue, Yibing Zhan 等AAAI 2025 · 被引用 1 次
- Kernel Functional OptimisationArun Kumar Anjanapura Venkatesh, Alistair Shilton, Santu Rana, Sunil Gupta 等NeurIPS 2021 · 被引用 6 次
