A variational approximate posterior for the deep Wishart process
Sebastian W. Ober, Laurence Aitchison
摘要
Recent work introduced deep kernel processes as an entirely kernel-based alternative to NNs (Aitchison et al. 2020) . Deep kernel processes flexibly learn good top-layer representations by alternately sampling the kernel from a distribution over positive semi-definite matrices and performing nonlinear transformations. A particular deep kernel process, the deep Wishart process (DWP), is of particular interest because its prior can be made equivalent to deep Gaussian process (DGP) priors for kernels that can be expressed entirely in terms of Gram matrices. However, inference in DWPs has not yet been possible due to the lack of sufficiently flexible distributions over positive semi-definite matrices. Here, we give a novel approach to obtaining flexible distributions over positive semi-definite matrices by generalising the Bartlett decomposition of the Wishart probability density. We use this new distribution to develop an approximate posterior for the DWP that includes dependency across layers. We develop a doubly-stochastic inducing-point inference scheme for the DWP and show experimentally that inference in the DWP can improve performance over doing inference in a DGP with the equivalent prior.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- A theory of representation learning gives a deep generalisation of kernel methodsAdam X. Yang, Maxime Robeyns, Edward Milsom, Ben Anson 等ICML 2023 · 被引用 15 次
- Stochastic Kernel Regularisation Improves Generalisation in Deep Kernel MachinesEdward Milsom, Ben Anson, Laurence AitchisonNeurIPS 2024 · 被引用 1 次
它引用的顶会 Paper3
- Global inducing point variational posteriors for Bayesian neural networks and deep Gaussian processesSebastian W. Ober, Laurence AitchisonICML 2021 · 被引用 65 次
- Why bigger is not always better: on finite and infinite neural networksLaurence AitchisonICML 2020 · 被引用 59 次
- Stochastic Differential Equations with Variational Wishart DiffusionsMartin Jørgensen, Marc Peter Deisenroth, Hugh SalimbeniICML 2020 · 被引用 8 次
相关 Paper
- Deep Kernel ProcessesLaurence Aitchison, Adam X. Yang, Sebastian W. OberICML 2021 · 被引用 44 次
- Deep Kernel Posterior Learning under Infinite Variance Prior WeightsJorge Loría, Anindya BhadraICLR 2025
- Bayesian Deep Ensembles via the Neural Tangent KernelBobby He, Balaji Lakshminarayanan, Yee Whye TehNeurIPS 2020 · 被引用 136 次
- Convolutional Deep Kernel MachinesEdward Milsom, Ben Anson, Laurence AitchisonICLR 2024 · 被引用 6 次
- On Neural Networks as Infinite Tree-Structured Probabilistic Graphical ModelsBoyao Li, Alexander Thomson, Houssam Nassif, Matthew Engelhard 等NeurIPS 2024 · 被引用 2 次
