Masked Bayesian Neural Networks : Theoretical Guarantee and its Posterior Inference
Insung Kong, Dongyoon Yang, Jongjin Lee, Ilsang Ohn, Gyuseung Baek, Yongdai Kim
Abstract
Bayesian approaches for learning deep neural networks (BNN) have been received much attention and successfully applied to various applications. Particularly, BNNs have the merit of having better generalization ability as well as better uncertainty quantification. For the success of BNN, search an appropriate architecture of the neural networks is an important task, and various algorithms to find good sparse neural networks have been proposed. In this paper, we propose a new node-sparse BNN model which has good theoretical properties and is computationally feasible. We prove that the posterior concentration rate to the true model is near minimax optimal and adaptive to the smoothness of the true model. In particular the adaptiveness is the first of its kind for node-sparse BNNs. In addition, we develop a novel MCMC algorithm which makes the Bayesian inference of the node-sparse BNN model feasible in practice.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Training Bayesian Neural Networks with Sparse Subspace Variational InferenceJunbo Li, Zichen Miao, Qiang Qiu, Ruqi ZhangICLR 2024 · 12 citations
- BayesTune: Bayesian Sparse Deep Model Fine-tuningMinyoung Kim, Timothy M. HospedalesNeurIPS 2023 · 10 citations
- Posterior Contraction for Sparse Neural Networks in Besov Spaces with Intrinsic DimensionalityKyeongwon Lee, Lizhen Lin, Jaewoo Park, Seonghyun JeongNeurIPS 2025 · 4 citations
- Knowledge Distillation of Uncertainty using Deep Latent Factor ModelSehyun Park, Jongjin Lee, Yunseop Shin, Ilsang Ohn et al.NeurIPS 2025 · 2 citations
- Bayesian Neural Networks for Functional ANOVA ModelSeokhun Park, Choeun Kim, Jihu Lee, Yunseop Shin et al.ICLR 2026
Builds on10
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 845 citations
- What Are Bayesian Neural Network Posteriors Really Like?Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman, Andrew Gordon WilsonICML 2021 · 458 citations
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski et al.ICML 2020 · 409 citations
- Oops I Took A Gradient: Scalable Sampling for Discrete DistributionsWill Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud et al.ICML 2021 · 113 citations
Related papers
- Convergence Rates of Variational Inference in Sparse Deep LearningBadr-Eddine Chérief-AbdellatifICML 2020 · 43 citations
- Efficient Variational Inference for Sparse Deep Learning with Theoretical GuaranteeJincheng Bai, Qifan Song, Guang ChengNeurIPS 2020 · 55 citations
- Flat Seeking Bayesian Neural NetworksVan-Anh Nguyen, Tung-Long Vuong, Hoang Phan, Thanh-Toan Do et al.NeurIPS 2023 · 14 citations
- Sparse Deep Learning: A New Framework Immune to Local Traps and MiscalibrationYan Sun, Wenjun Xiong, Faming LiangNeurIPS 2021 · 11 citations
- Asymptotic Properties for Bayesian Neural Network in Besov SpaceKyeongwon Lee, Jaeyong LeeNeurIPS 2022 · 8 citations
