Lune

ICML2021Top-tier venue

High-Dimensional Gaussian Process Inference with Derivatives

Filip de Roos, Alexandra Gessner, Philipp Hennig

2021Year
24Citations
6Top-tier citations

Abstract

Although it is widely known that Gaussian processes can be conditioned on observations of the gradient, this functionality is of limited use due to the prohibitive computational cost of O(N3D3)\mathcal{O}(N^3 D^3) in data points NN and dimension DD. The dilemma of gradient observations is that a single one of them comes at the same cost as DD independent function evaluations, so the latter are often preferred. Careful scrutiny reveals, however, that derivative observations give rise to highly structured kernel Gram matrices for very general classes of kernels (inter alia, stationary kernels). We show that in the low-data regime N<DN<D, the Gram matrix can be decomposed in a manner that reduces the cost of inference to O(N2D+(N2)3)\mathcal{O}(N^2D + (N^2)^3) (i.e., linear in the number of dimensions) and, in special cases, to O(N2D+N3)\mathcal{O}(N^2D + N^3). This reduction in complexity opens up new use-cases for inference with gradients especially in the high-dimensional regime, where the information-to-cost ratio of gradient observations significantly increases. We demonstrate this potential in a variety of tasks relevant for machine learning, such as optimization and Hamiltonian Monte Carlo with predictive gradients.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext caf38627-afa5-4584-b441-3bd506219d69

Cited by top-tier papers6

Ask how each one uses it

Builds on2

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines