High-Dimensional Gaussian Process Inference with Derivatives
Filip de Roos, Alexandra Gessner, Philipp Hennig
摘要
Although it is widely known that Gaussian processes can be conditioned on observations of the gradient, this functionality is of limited use due to the prohibitive computational cost of in data points and dimension . The dilemma of gradient observations is that a single one of them comes at the same cost as independent function evaluations, so the latter are often preferred. Careful scrutiny reveals, however, that derivative observations give rise to highly structured kernel Gram matrices for very general classes of kernels (inter alia, stationary kernels). We show that in the low-data regime , the Gram matrix can be decomposed in a manner that reduces the cost of inference to (i.e., linear in the number of dimensions) and, in special cases, to . This reduction in complexity opens up new use-cases for inference with gradients especially in the high-dimensional regime, where the information-to-cost ratio of gradient observations significantly increases. We demonstrate this potential in a variety of tasks relevant for machine learning, such as optimization and Hamiltonian Monte Carlo with predictive gradients.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Unexpected Improvements to Expected Improvement for Bayesian OptimizationSebastian Ament, Samuel Daulton, David Eriksson, Maximilian Balandat 等NeurIPS 2023 · 被引用 280 次
- Scaling Gaussian Processes with Derivative Information Using Variational InferenceMisha Padidar, Xinran Zhu, Leo Huang, Jacob R. Gardner 等NeurIPS 2021 · 被引用 28 次
- Scalable First-Order Bayesian Optimization via Structured Automatic DifferentiationSebastian E. Ament, Carla P. GomesICML 2022 · 被引用 12 次
- Monotonicity and Double Descent in Uncertainty Estimation with Gaussian ProcessesLiam Hodgkinson, Christopher van der Heide, Fred Roosta, Michael W. MahoneyICML 2023 · 被引用 9 次
- BayeSQP: Bayesian Optimization through Sequential Quadratic ProgrammingPaul Brunzema, Sebastian TrimpeNeurIPS 2025 · 被引用 7 次
它引用的顶会 Paper2
相关 Paper
- Scalable Gaussian Processes with Latent Kronecker StructureJihao Andreas Lin, Sebastian Ament, Maximilian Balandat, David Eriksson 等ICML 2025
- The Price of Linear Time: Error Analysis of Structured Kernel InterpolationAlexander Moreno, Justin Xiao, Jonathan MeiICML 2025
- Sampling from Gaussian Process Posteriors using Stochastic Gradient DescentJihao Andreas Lin, Javier Antorán, Shreyas Padhy, David Janz 等NeurIPS 2023 · 被引用 34 次
- KernelMatmul: Scaling Gaussian Processes to Large Time SeriesTilman Hoffbauer, Holger H. Hoos, Jakob BossekAAAI 2025
- Variational Sparse Inverse Cholesky Approximation for Latent Gaussian Processes via Double Kullback-Leibler MinimizationJian Cao, Myeongjong Kang, Felix Jimenez, Huiyan Sang 等ICML 2023 · 被引用 12 次
