Efficient Low Rank Gaussian Variational Inference for Neural Networks
Marcin Tomczak, Siddharth Swaroop, Richard E. Turner
摘要
Bayesian neural networks are enjoying a renaissance driven in part by recent advances in variational inference (VI). The most common form of VI employs a fully factorized or mean-field distribution, but this is known to suffer from several pathologies, especially as we expect posterior distributions with highly correlated parameters. Current algorithms that capture these correlations with a Gaussian approximating family are difficult to scale to large models due to computational costs and high variance of gradient updates. By using a new form of the reparametrization trick, we derive a computationally efficient algorithm for performing VI with a Gaussian family with a low-rank plus diagonal covariance structure. We scale to deep feed-forward and convolutional architectures. We find that adding low-rank terms to parametrized diagonal covariance does not improve predictive performance except on small networks, but low-rank terms added to a constant diagonal covariance improves performance on small and large-scale network architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- BLoB: Bayesian Low-Rank Adaptation by Backpropagation for Large Language ModelsYibin Wang, Haizhou Shi, Ligong Han, Dimitris N. Metaxas 等NeurIPS 2024 · 被引用 63 次
- The Empirical Impact of Neural Parameter Symmetries, or Lack ThereofDerek Lim, Theo (Moe) Putterman, Robin Walters, Haggai Maron 等NeurIPS 2024 · 被引用 25 次
- Accurate Node Feature Estimation with Structured Variational Graph AutoencoderJaemin Yoo, Hyunsik Jeon, Jinhong Jung, U KangKDD 2022 · 被引用 24 次
- Bayesian Online Natural Gradient (BONG)Matt Jones, Peter G. Chang, Kevin P. MurphyNeurIPS 2024 · 被引用 20 次
- Towards Understanding the Dynamics of Gaussian-Stein Variational Gradient DescentTianle Liu, Promit Ghosal, Krishnakumar Balasubramanian, Natesh S. PillaiNeurIPS 2023 · 被引用 19 次
它引用的顶会 Paper1
相关 Paper
- VIKING: Deep variational inference with stochastic projectionsSamuel Matthiesen, Hrittik Roy, Nicholas Krämer, Yevgen Zainchkovskyy 等NeurIPS 2025 · 被引用 3 次
- Liberty or Depth: Deep Bayesian Neural Nets Do Not Need Complex Weight Posterior ApproximationsSebastian Farquhar, Lewis Smith, Yarin GalNeurIPS 2020 · 被引用 47 次
- Provably Scalable Black-Box Variational Inference with Structured Variational FamiliesJoohwan Ko, Kyurae Kim, Woochang Kim, Jacob R. GardnerICML 2024 · 被引用 6 次
- Walsh-Hadamard Variational Inference for Bayesian Deep LearningSimone Rossi, Sébastien Marmin, Maurizio FilipponeNeurIPS 2020 · 被引用 17 次
- Dissecting Non-Vacuous Generalization Bounds based on the Mean-Field ApproximationKonstantinos PitasICML 2020 · 被引用 8 次
