Amortized Variational Deep Kernel Learning
Alan L. S. Matias, César Lincoln C. Mattos, João Paulo Pordeus Gomes, Diego Mesquita
摘要
Deep kernel learning (DKL) marries the uncertainty quantification of Gaussian processes (GPs) and the representational power of deep neural networks. However, training DKL is challenging and often leads to overfitting. Most notably, DKL often learns "non-local" kernels -incurring spurious correlations. To remedy this issue, we propose using amortized inducing points and a parameter-sharing scheme, which ties together the amortization and DKL networks. This design imposes an explicit dependency between the ELBO's model fit and capacity terms. In turn, this prevents the former from dominating the optimization procedure and incurring the aforementioned spurious correlations. Extensive experiments show that our resulting method, amortized varitional DKL (AVDKL), i) consistently outperforms DKL and standard GPs for tabular data; ii) achieves significantly higher accuracy than DKL in node classification tasks; and iii) leads to substantially better accuracy and negative loglikelihood than DKL on CIFAR100.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Transformed Latent Variable Multi-Output Gaussian ProcessesXiaoyu Jiang, Xinxing Shi, Sokratia Georgaka, Magnus Rattray 等ICML 2026
- GPan-LoRA: Gaussian Process Amortized Networks for Bayesian Low-Rank Adaptation in Large Language ModelsWeifeng Zhang, Wenyuan Zhao, Amir Hossein Rahmati, Yucheng Wang 等ICML 2026
它引用的顶会 Paper8
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- Neural Tangents: Fast and Easy Infinite Neural Networks in PythonRoman Novak, Lechao Xiao, Jiri Hron, Jaehoon Lee 等ICLR 2020 · 被引用 254 次
- Infinite attention: NNGP and NTK for deep attention networksJiri Hron, Yasaman Bahri, Jascha Sohl-Dickstein, Roman NovakICML 2020 · 被引用 147 次
- Personalized Federated Learning With Gaussian ProcessesIdan Achituve, Aviv Shamsian, Aviv Navon, Gal Chechik 等NeurIPS 2021 · 被引用 137 次
- Fast Finite Width Neural Tangent KernelRoman Novak, Jascha Sohl-Dickstein, Samuel S. SchoenholzICML 2022 · 被引用 72 次
相关 Paper
- Longitudinal Deep Kernel Gaussian Process RegressionJunjie Liang, Yanting Wu, Dongkuan Xu, Vasant G. HonavarAAAI 2021 · 被引用 9 次
- Input Dependent Sparse Gaussian ProcessesBahram Jafrasteh, Carlos Villacampa-Calvo, Daniel Hernández-LobatoICML 2022 · 被引用 7 次
- Global inducing point variational posteriors for Bayesian neural networks and deep Gaussian processesSebastian W. Ober, Laurence AitchisonICML 2021 · 被引用 65 次
- A theory of representation learning gives a deep generalisation of kernel methodsAdam X. Yang, Maxime Robeyns, Edward Milsom, Ben Anson 等ICML 2023 · 被引用 15 次
- Fully Bayesian Autoencoders with Latent Sparse Gaussian ProcessesBa-Hien Tran, Babak Shahbaba, Stephan Mandt, Maurizio FilipponeICML 2023 · 被引用 9 次
