Sparse Uncertainty Representation in Deep Learning with Inducing Weights
Hippolyt Ritter, Martin Kukla, Cheng Zhang, Yingzhen Li
摘要
Bayesian neural networks and deep ensembles represent two modern paradigms of uncertainty quantification in deep learning. Yet these approaches struggle to scale mainly due to memory inefficiency issues, since they require parameter storage several times higher than their deterministic counterparts. To address this, we augment the weight matrix of each layer with a small number of inducing weights, thereby projecting the uncertainty quantification into such low dimensional spaces. We further extend Matheron's conditional Gaussian sampling rule to enable fast weight sampling, which enables our inference method to maintain reasonable run-time as compared with ensembles. Importantly, our approach achieves competitive performance to the state-of-the-art in prediction and uncertainty estimation tasks with fully connected neural networks and ResNets, while reducing the parameter size to of that of a neural network.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Calibrating Multimodal LearningHuan Ma, Qingyang Zhang, Changqing Zhang, Bingzhe Wu 等ICML 2023 · 被引用 42 次
- Structured Stochastic Gradient MCMCAntonios Alexos, Alex J. Boyd, Stephan MandtICML 2022 · 被引用 14 次
- Function-space Inference with Sparse Implicit ProcessesSimón Rodríguez Santana, Bryan Zaldivar, Daniel Hernández-LobatoICML 2022 · 被引用 13 次
- Training Bayesian Neural Networks with Sparse Subspace Variational InferenceJunbo Li, Zichen Miao, Qiang Qiu, Ruqi ZhangICLR 2024 · 被引用 12 次
- Self-Attention through Kernel-Eigen Pair Sparse Variational Gaussian ProcessesYingyi Chen, Qinghua Tao, Francesco Tonin, Johan A. K. SuykensICML 2024 · 被引用 4 次
它引用的顶会 Paper8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 被引用 569 次
- Cyclical Stochastic Gradient MCMC for Bayesian Deep LearningRuqi Zhang, Chunyuan Li, Jianyi Zhang, Changyou Chen 等ICLR 2020 · 被引用 292 次
- Efficient and Scalable Bayesian Neural Nets with Rank-1 FactorsMichael Dusenberry, Ghassen Jerfel, Yeming Wen, Yi-An Ma 等ICML 2020 · 被引用 239 次
- Efficiently sampling functions from Gaussian process posteriorsJames T. Wilson, Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky 等ICML 2020 · 被引用 186 次
相关 Paper
- Connecting the Dots: Is Mode-Connectedness the Key to Feasible Sample-Based Inference in Bayesian Neural Networks?Emanuel Sommer, Lisa Wimmer, Theodore Papamarkou, Ludwig Bothmann 等ICML 2024
- Uncertainty Quantification with the Empirical Neural Tangent KernelJoseph Wilson, Chris van der Heide, Liam Hodgkinson, Fred RoostaNeurIPS 2025 · 被引用 11 次
- Collapsed Inference for Bayesian Deep LearningZhe Zeng, Guy Van den BroeckNeurIPS 2023 · 被引用 10 次
- Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution ShiftsHa Manh Bui, Anqi LiuICML 2024 · 被引用 12 次
- Implicit Neural Representation Inference for Low-Dimensional Bayesian Deep LearningPanagiotis Dimitrakopoulos, Giorgos Sfikas, Christophoros NikouICLR 2024
