Unrolled denoising networks provably learn to perform optimal Bayesian inference
Aayush Karan, Kulin Shah, Sitan Chen, Yonina C. Eldar
Abstract
Much of Bayesian inference centers around the design of estimators for inverse problems which are optimal assuming the data comes from a known prior. But what do these optimality guarantees mean if the prior is unknown? In recent years, algorithm unrolling has emerged as deep learning’s answer to this age-old question: design a neural network whose layers can in principle simulate iterations of inference algorithms and train on data generated by the unknown prior. Despite its empirical success, however, it has remained unclear whether this method can provably recover the performance of its optimal, prior-aware counterparts. In this work, we prove the first rigorous learning guarantees for neural networks based on unrolling approximate message passing (AMP). For compressed sensing, we prove that when trained on data drawn from a product prior, the layers of the network approximately converge to the same denoisers used in Bayes AMP. We also provide extensive numerical experiments for compressed sensing and rank-one matrix estimation demonstrating the advantages of our unrolled architecture – in addition to being able to obliviously adapt to general priors, it exhibits improvements over Bayes AMP in more general settings of low dimensions, non-Gaussian designs, and non-product priors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Time-Embedded Algorithm Unrolling for Computational MRIJunno Yun, Yasar Utku Alçalar, Mehmet AkçakayaNeurIPS 2025 · 9 citations
- How many measurements are enough? Bayesian recovery in inverse problems with general distributionsBen Adcock, Zi Yuan (Nick) HuangNeurIPS 2025 · 2 citations
Builds on8
- Convergence for score-based generative modeling with polynomial complexityHolden Lee, Jianfeng Lu, Yixin TanNeurIPS 2022 · 221 citations
- On the linearity of large non-linear models: when and why the tangent kernel is constantChaoyue Liu, Libin Zhu, Mikhail BelkinNeurIPS 2020 · 183 citations
- Learning Mixtures of Gaussians Using the DDPM ObjectiveKulin Shah, Sitan Chen, Adam R. KlivansNeurIPS 2023 · 69 citations
- Hyperparameter Tuning is All You Need for LISTAXiaohan Chen, Jialin Liu, Zhangyang Wang, Wotao YinNeurIPS 2021 · 40 citations
- The price of ignorance: how much does it cost to forget noise structure in low-rank matrix estimation?Jean Barbier, TianQi Hou, Marco Mondelli, Manuel SáenzNeurIPS 2022 · 25 citations
Related papers
- PCA Initialization for Approximate Message Passing in Rotationally Invariant ModelsMarco Mondelli, Ramji VenkataramananNeurIPS 2021 · 23 citations
- What's in a Prior? Learned Proximal Networks for Inverse ProblemsZhenghan Fang, Sam Buchanan, Jeremias SulamICLR 2024 · 27 citations
- Bayes-optimal learning of an extensive-width neural network from quadratically many samplesAntoine Maillard, Emanuele Troiani, Simon Martin, Florent Krzakala et al.NeurIPS 2024 · 26 citations
- Multi-layer State Evolution Under Random Convolutional DesignMax Daniels, Cédric Gerbelot, Florent Krzakala, Lenka ZdeborováNeurIPS 2022
- Probabilistic Unrolling: Scalable, Inverse-Free Maximum Likelihood Estimation for Latent Gaussian ModelsAlexander Lin, Bahareh Tolooshams, Yves F. Atchadé, Demba E. BaICML 2023 · 1 citation
