Scaling Gaussian Processes with Derivative Information Using Variational Inference
Misha Padidar, Xinran Zhu, Leo Huang, Jacob R. Gardner, David Bindel
Abstract
Gaussian processes with derivative information are useful in many settings where derivative information is available, including numerous Bayesian optimization and regression tasks that arise in the natural sciences. Incorporating derivative observations, however, comes with a dominating computational cost when training on points in input dimensions. This is intractable for even moderately sized problems. While recent work has addressed this intractability in the low- setting, the high-, high- setting is still unexplored and of great value, particularly as machine learning problems increasingly become high dimensional. In this paper, we introduce methods to achieve fully scalable Gaussian process regression with derivatives using variational inference. Analogous to the use of inducing values to sparsify the labels of a training set, we introduce the concept of inducing directional derivatives to sparsify the partial derivative information of a training set. This enables us to construct a variational posterior that incorporates derivative information but whose size depends neither on the full dataset size nor the full dimensionality . We demonstrate the full scalability of our approach on a variety of tasks, ranging from a high dimensional stellarator fusion regression task to training graph convolutional neural networks on Pubmed using Bayesian optimization. Surprisingly, we find that our approach can improve regression performance even in settings where only label data is available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4cd2e7a4-5c2c-4bb2-b429-632af3b0b811Cited by top-tier papers5
- Scalable First-Order Bayesian Optimization via Structured Automatic DifferentiationSebastian E. Ament, Carla P. GomesICML 2022 · 12 citations
- Physics-Informed Variational State-Space Gaussian ProcessesOliver Hamelijnck, Arno Solin, Theodoros DamoulasNeurIPS 2024 · 12 citations
- Random Function DescentFelix Benning, Leif DöringNeurIPS 2024 · 1 citation
- Efficient Hyperparameter Optimization with Adaptive Fidelity IdentificationJiantong Jiang, Zeyi Wen, Atif Bin Mansoor, Ajmal MianCVPR 2024
- Monotonic Variational Gaussian Process for Efficient Data CollectionDonghyun Lee, Young Myoung KoICML 2026
Builds on3
- Parametric Gaussian Process RegressorsMartin Jankowiak, Geoff Pleiss, Jacob R. GardnerICML 2020 · 82 citations
- Fast Matrix Square Roots with Applications to Gaussian Processes and Bayesian OptimizationGeoff Pleiss, Martin Jankowiak, David Eriksson, Anil Damle et al.NeurIPS 2020 · 49 citations
- High-Dimensional Gaussian Process Inference with DerivativesFilip de Roos, Alexandra Gessner, Philipp HennigICML 2021 · 24 citations
Related papers
- Input Dependent Sparse Gaussian ProcessesBahram Jafrasteh, Carlos Villacampa-Calvo, Daniel Hernández-LobatoICML 2022 · 7 citations
- Implicit Manifold Gaussian Process RegressionBernardo Fichera, Slava Borovitskiy, Andreas Krause, Aude Gemma BillardNeurIPS 2023 · 10 citations
- Deep Random Features for Scalable Interpolation of Spatiotemporal DataWeibin Chen, Azhir Mahmood, Michel Tsamados, So TakaoICLR 2025
- Sparse within Sparse Gaussian Processes using Neighbor InformationGia-Lac Tran, Dimitrios Milios, Pietro Michiardi, Maurizio FilipponeICML 2021 · 19 citations
- Bezier Gaussian Processes for Tall and Wide DataMartin Jørgensen, Michael A. OsborneNeurIPS 2022 · 2 citations
