Learning Expressive Priors for Generalization and Uncertainty Estimation in Neural Networks
Dominik Schnaus, Jongseok Lee, Daniel Cremers, Rudolph Triebel
Abstract
In this work, we propose a novel prior learning method for advancing generalization and uncertainty estimation in deep neural networks. The key idea is to exploit scalable and structured posteriors of neural networks as informative priors with generalization guarantees. Our learned priors provide expressive probabilistic representations at large scale, like Bayesian counterparts of pretrained models on ImageNet, and further produce non-vacuous generalization bounds. We also extend this idea to a continual learning framework, where the favorable properties of our priors are desirable. Major enablers are our technical contributions: (1) the sums-of-Kronecker-product computations, and (2) the derivations and optimizations of tractable objectives that lead to improved generalization bounds. Empirically, we exhaustively show the effectiveness of this method for uncertainty estimation and generalization. * Equal contribution 1 TU Munich 2 MCML 3 DLR 4 KIT. Correspondence to: Dominik Schnaus
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6e1901fb-4bfd-47ca-aaef-628453bd9067Cited by top-tier papers2
- General Uncertainty Estimation with Delta VariancesSimon Schmitt, John Shawe-Taylor, Hado van HasseltAAAI 2025 · 3 citations
- Do Bayesian Neural Networks Actually Behave Like Bayesian Models?Gábor Pituk, Vik Shirvaikar, Tom RainforthICML 2025
Builds on16
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 845 citations
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 569 citations
- Laplace Redux - Effortless Bayesian Deep LearningErik A. Daxberger, Agustinus Kristiadi, Alexander Immer, Runa Eschenhagen et al.NeurIPS 2021 · 508 citations
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski et al.ICML 2020 · 409 citations
- Uncertainty-guided Continual Learning with Bayesian Neural NetworksSayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, Marcus RohrbachICLR 2020 · 211 citations
Related papers
- FSP-Laplace: Function-Space Priors for the Laplace Approximation in Bayesian Deep LearningTristan Cinquin, Marvin Pförtner, Vincent Fortuin, Philipp Hennig et al.NeurIPS 2024 · 15 citations
- Bayesian Neural Scaling Law Extrapolation with Prior-Data Fitted NetworksDongwoo Lee, Dong Bok Lee, Steven Adriaensen, Juho Lee et al.ICML 2025
- MARS: Meta-learning as Score Matching in the Function SpaceKrunoslav Lehman Pavasovic, Jonas Rothfuss, Andreas KrauseICLR 2023 · 1 citation
- Bayesian Structural Adaptation for Continual LearningAbhishek Kumar, Sunabha Chatterjee, Piyush RaiICML 2021 · 7 citations
- Flat Seeking Bayesian Neural NetworksVan-Anh Nguyen, Tung-Long Vuong, Hoang Phan, Thanh-Toan Do et al.NeurIPS 2023 · 14 citations
