U-Nets as Belief Propagation: Efficient Classification, Denoising, and Diffusion in Generative Hierarchical Models
Song Mei
摘要
U-Nets are among the most widely used architectures in computer vision, renowned for their exceptional performance in applications such as image segmentation, denoising, and diffusion modeling. However, a theoretical explanation of the U-Net architecture design has not yet been fully established. This paper introduces a novel interpretation of the U-Net architecture by studying certain generative hierarchical models, which are tree-structured graphical models extensively utilized in both language and image domains. With their encoder-decoder structure, long skip connections, and pooling and upsampling layers, we demonstrate how U-Nets can naturally implement the belief propagation denoising algorithm in such generative hierarchical models, thereby efficiently approximating the denoising functions. This leads to an efficient sample complexity bound for learning the denoising function using U-Nets within these models. Additionally, we discuss the broader implications of these findings for diffusion models in generative hierarchical models. We also demonstrate that the conventional architecture of convolutional neural networks (ConvNets) is ideally suited for classification tasks within these models. This offers a unified view of the roles of ConvNets and U-Nets, highlighting the versatility of generative hierarchical models in modeling complex data distributions across language and image domains.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Towards a theory of how the structure of language is acquired by deep neural networksFrancesco Cagnetta, Matthieu WyartNeurIPS 2024 · 被引用 33 次
- A solvable model of learning generative diffusion: theory and insightsHugo Cui, Cengiz Pehlevan, Yue M. LuNeurIPS 2025 · 被引用 11 次
- Time-Embedded Algorithm Unrolling for Computational MRIJunno Yun, Yasar Utku Alçalar, Mehmet AkçakayaNeurIPS 2025 · 被引用 9 次
- Unrolled denoising networks provably learn to perform optimal Bayesian inferenceAayush Karan, Kulin Shah, Sitan Chen, Yonina C. EldarNeurIPS 2024 · 被引用 5 次
- Transformers as Unsupervised Learning Algorithms: A study on Gaussian MixturesZhiheng Chen, Ruofan Wu, Guanhua FangICLR 2026 · 被引用 2 次
它引用的顶会 Paper28
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
- UCTransNet: Rethinking the Skip Connections in U-Net from a Channel-Wise Perspective with TransformerHaonan Wang, Peng Cao, Jiaqi Wang, Osmar R. ZaïaneAAAI 2022 · 被引用 1,144 次
- Improved Analysis of Score-based Generative Modeling: User-Friendly Bounds under Minimal Smoothness AssumptionsHongrui Chen, Holden Lee, Jianfeng LuICML 2023 · 被引用 212 次
- The probability flow ODE is provably fastSitan Chen, Sinho Chewi, Holden Lee, Yuanzhi Li 等NeurIPS 2023 · 被引用 179 次
相关 Paper
- A Unified Framework for U-Net Design and AnalysisChristopher Williams, Fabian Falck, George Deligiannidis, Chris C. Holmes 等NeurIPS 2023 · 被引用 79 次
- A Multi-Resolution Framework for U-Nets with Applications to Hierarchical VAEsFabian Falck, Christopher Williams, Dominic Danks, George Deligiannidis 等NeurIPS 2022 · 被引用 11 次
- Neural Residual Diffusion Models for Deep Scalable Vision GenerationZhiyuan Ma, Liangliang Zhao, Biqing Qi, Bowen ZhouNeurIPS 2024 · 被引用 15 次
- Modeling Structure with Undirected Neural NetworksTsvetomila Mihaylova, Vlad Niculae, André F. T. MartinsICML 2022 · 被引用 1 次
- How Compositional Generalization and Creativity Improve as Diffusion Models are TrainedAlessandro Favero, Antonio Sclocchi, Francesco Cagnetta, Pascal Frossard 等ICML 2025
