Variational Inference for Infinitely Deep Neural Networks
Achille Nazaret, David M. Blei
Abstract
We introduce the unbounded depth neural network (UDN), an infinitely deep probabilistic model that adapts its complexity to the training data. The UDN contains an infinite sequence of hidden layers and places an unbounded prior on a truncation L, the layer from which it produces its data. Given a dataset of observations, the posterior UDN provides a conditional distribution of both the parameters of the infinite neural network and its truncation. We develop a novel variational inference algorithm to approximate this posterior, optimizing a distribution of the neural network weights and of the truncation depth L, and without any upper limit on L. To this end, the variational family has a special structure: it models neural network weights of arbitrary depth, and it dynamically creates or removes free variational parameters as its distribution of the truncation is optimized. (Unlike heuristic approaches to model search, it is solely through gradient-based optimization that this algorithm explores the space of truncations.) We study the UDN on real and synthetic data. We find that the UDN adapts its posterior depth to the dataset complexity; it outperforms standard neural networks of similar computational complexity; and it outperforms other approaches to infinite-depth neural networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e86348d3-a4ce-4999-8768-05e75720cd4cCited by top-tier papers4
- Bayesian Adaptation of Network Depth and Width for Continual LearningJeevan Thapa, Rui LiICML 2024 · 7 citations
- Adaptive Width Neural NetworksFederico Errica, Henrik Christiansen, Viktor Zaverkin, Mathias Niepert et al.ICLR 2026 · 6 citations
- Enhancing Diversity in Bayesian Deep Learning via Hyperspherical Energy Minimization of CKADavid Smerkous, Qinxun Bai, Fuxin LiNeurIPS 2024 · 3 citations
- Adaptive Message Passing: A General Framework to Mitigate Oversmoothing, Oversquashing, and UnderreachingFederico Errica, Henrik Christiansen, Viktor Zaverkin, Takashi Maruyama et al.ICML 2025
Builds on4
- Multiscale Deep Equilibrium ModelsShaojie Bai, Vladlen Koltun, J. Zico KolterNeurIPS 2020 · 272 citations
- Neural Tangents: Fast and Easy Infinite Neural Networks in PythonRoman Novak, Lechao Xiao, Jiri Hron, Jaehoon Lee et al.ICLR 2020 · 254 citations
- Depth Uncertainty in Neural NetworksJavier Antorán, James Urquhart Allingham, José Miguel Hernández-LobatoNeurIPS 2020 · 121 citations
- The Limitations of Large Width in Neural Networks: A Deep Gaussian Process PerspectiveGeoff Pleiss, John P. CunninghamNeurIPS 2021 · 35 citations
Related papers
- Partially Stochastic Infinitely Deep Bayesian Neural NetworksSergio Calvo-Ordoñez, Matthieu Meunier, Francesco Piatti, Yuantao ShiICML 2024 · 7 citations
- Specifying Weight Priors in Bayesian Deep Neural Networks with Empirical BayesRanganath Krishnan, Mahesh Subedar, Omesh TickooAAAI 2020 · 65 citations
- Masked Bayesian Neural Networks : Theoretical Guarantee and its Posterior InferenceInsung Kong, Dongyoon Yang, Jongjin Lee, Ilsang Ohn et al.ICML 2023 · 8 citations
- Liberty or Depth: Deep Bayesian Neural Nets Do Not Need Complex Weight Posterior ApproximationsSebastian Farquhar, Lewis Smith, Yarin GalNeurIPS 2020 · 47 citations
- Effective Estimation of Deep Generative Language ModelsTom Pelsmaeker, Wilker AzizACL 2020 · 5 citations
