A Rigorous Link between Deep Ensembles and (Variational) Bayesian Methods
Veit David Wild, Sahra Ghalebikesabi, Dino Sejdinovic, Jeremias Knoblauch
摘要
We establish the first mathematically rigorous link between Bayesian, variational Bayesian, and ensemble methods. A key step towards this it to reformulate the non-convex optimisation problem typically encountered in deep learning as a convex optimisation in the space of probability measures. On a technical level, our contribution amounts to studying generalised variational inference through the lense of Wasserstein gradient flows. The result is a unified theory of various seemingly disconnected approaches that are commonly used for uncertainty quantification in deep learning -- including deep ensembles and (variational) Bayesian methods. This offers a fresh perspective on the reasons behind the success of deep ensembles over procedures based on parameterised variational inference, and allows the derivation of new ensembling schemes with convergence guarantees. We showcase this by proposing a family of interacting deep ensembles with direct parallels to the interactions of particle systems in thermodynamics, and use our theory to prove the convergence of these algorithms to a well-defined global minimiser on the space of probability measures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Bayesian Uncertainty for Gradient Aggregation in Multi-Task LearningIdan Achituve, Idit Diamant, Arnon Netzer, Gal Chechik 等ICML 2024 · 被引用 14 次
- TabMGP: Martingale Posterior with TabPFNKenyon Ng, Edwin Fong, David Frazier, Jeremias Knoblauch 等ICML 2026 · 被引用 8 次
- Variational Deep Learning via Implicit RegularizationJonathan Wenger, Beau Coker, Juraj Marusic, John Patrick CunninghamICLR 2026 · 被引用 1 次
- Bayesian Low-Rank Learning (Bella): A Practical Approach to Bayesian Neural NetworksBao Gia Doan, Afshar Shamsi, Xiao-Yu Guo, Arash Mohammadi 等AAAI 2025 · 被引用 1 次
- Robust Bayesian Optimisation with Unbounded CorruptionsAbdelhamid Ezzerg, Ilija Bogunovic, Jeremias KnoblauchICML 2026 · 被引用 1 次
它引用的顶会 Paper10
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 被引用 845 次
- What Are Bayesian Neural Network Posteriors Really Like?Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman, Andrew Gordon WilsonICML 2021 · 被引用 458 次
- Repulsive Deep Ensembles are BayesianFrancesco D'Angelo, Vincent FortuinNeurIPS 2021 · 被引用 141 次
- Kernel Stein Discrepancy DescentAnna Korba, Pierre-Cyril Aubin-Frankowski, Szymon Majewski, Pierre AblinICML 2021 · 被引用 64 次
- KALE Flow: A Relaxed KL Gradient Flow for Probabilities with Disjoint SupportPierre Glaser, Michael Arbel, Arthur GrettonNeurIPS 2021 · 被引用 49 次
相关 Paper
- Is Epistemic Uncertainty Faithfully Represented by Evidential Deep Learning Methods?Mira Jürgens, Nis Meinert, Viktor Bengs, Eyke Hüllermeier 等ICML 2024 · 被引用 35 次
- Bayesian Posterior Approximation With Stochastic EnsemblesOleksandr Balabanov, Bernhard Mehlig, Hampus LinanderCVPR 2023
- Microcanonical Langevin Ensembles: Advancing the Sampling of Bayesian Neural NetworksEmanuel Sommer, Jakob Robnik, Giorgi Nozadze, Uros Seljak 等ICLR 2025
- Efficient and Scalable Bayesian Neural Nets with Rank-1 FactorsMichael Dusenberry, Ghassen Jerfel, Yeming Wen, Yi-An Ma 等ICML 2020 · 被引用 239 次
- Flat Seeking Bayesian Neural NetworksVan-Anh Nguyen, Tung-Long Vuong, Hoang Phan, Thanh-Toan Do 等NeurIPS 2023 · 被引用 14 次
