The k-tied Normal Distribution: A Compact Parameterization of Gaussian Mean Field Posteriors in Bayesian Neural Networks
Jakub Swiatkowski, Kevin Roth, Bastiaan S. Veeling, Linh Tran, Joshua V. Dillon, Jasper Snoek, Stephan Mandt, Tim Salimans, Rodolphe Jenatton, Sebastian Nowozin
摘要
Variational Bayesian Inference is a popular methodology for approximating posterior distributions over Bayesian neural network weights. Recent work developing this class of methods has explored ever richer parameterizations of the approximate posterior in the hope of improving performance. In contrast, here we share a curious experimental finding that suggests instead restricting the variational distribution to a more compact parameterization. For a variety of deep Bayesian neural networks trained using Gaussian mean-field variational inference, we find that the posterior standard deviations consistently exhibit strong low-rank structure after convergence. This means that by decomposing these variational parameters into a low-rank factorization, we can make our variational approximation more compact without decreasing the models' performance. Furthermore, we find that such factorized parameterizations improve the signal-to-noise ratio of stochastic gradient estimates of the variational lower bound, resulting in faster convergence.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Efficient and Scalable Bayesian Neural Nets with Rank-1 FactorsMichael Dusenberry, Ghassen Jerfel, Yeming Wen, Yi-An Ma 等ICML 2020 · 被引用 239 次
- Bayesian Neural Network Priors RevisitedVincent Fortuin, Adrià Garriga-Alonso, Sebastian W. Ober, Florian Wenzel 等ICLR 2022 · 被引用 162 次
- Repulsive Deep Ensembles are BayesianFrancesco D'Angelo, Vincent FortuinNeurIPS 2021 · 被引用 141 次
- Bayesian Deep Learning via Subnetwork InferenceErik A. Daxberger, Eric T. Nalisnick, James Urquhart Allingham, Javier Antorán 等ICML 2021 · 被引用 108 次
- Liberty or Depth: Deep Bayesian Neural Nets Do Not Need Complex Weight Posterior ApproximationsSebastian Farquhar, Lewis Smith, Yarin GalNeurIPS 2020 · 被引用 47 次
相关 Paper
- Efficient Low Rank Gaussian Variational Inference for Neural NetworksMarcin Tomczak, Siddharth Swaroop, Richard E. TurnerNeurIPS 2020 · 被引用 37 次
- Radial and Directional Posteriors for Bayesian Deep LearningChangYong Oh, Kamil Adamczewski, Mijung ParkAAAI 2020 · 被引用 6 次
- VIKING: Deep variational inference with stochastic projectionsSamuel Matthiesen, Hrittik Roy, Nicholas Krämer, Yevgen Zainchkovskyy 等NeurIPS 2025 · 被引用 3 次
- Collapsed Variational Bounds for Bayesian Neural NetworksMarcin Tomczak, Siddharth Swaroop, Andrew Y. K. Foong, Richard E. TurnerNeurIPS 2021 · 被引用 14 次
- Walsh-Hadamard Variational Inference for Bayesian Deep LearningSimone Rossi, Sébastien Marmin, Maurizio FilipponeNeurIPS 2020 · 被引用 17 次
