Shallow diffusion networks provably learn hidden low-dimensional structure
Nicholas Matthew Boffi, Arthur Jacot, Stephen Tu, Ingvar M. Ziemann
Abstract
Diffusion-based generative models provide a powerful framework for learning to sample from a complex target distribution. The remarkable empirical success of these models applied to high-dimensional signals, including images and video, stands in stark contrast to classical results highlighting the curse of dimensionality for distribution recovery. In this work, we take a step towards understanding this gap through a careful analysis of learning diffusion models over the Barron space of single layer neural networks. In particular, we show that these shallow models provably adapt to simple forms of low dimensional structure, thereby avoiding the curse of dimensionality. We combine our results with recent analyses of sampling with diffusion models to provide an end-to-end sample complexity bound for learning to sample from structured distributions. Importantly, our results do not require specialized architectures tailored to particular latent structures, and instead rely on the low-index structure of the Barron space to adapt to the underlying distribution. To implement this scheme in practice, the score function must be learned, and the reverse process must be discretized. Assuming access to a learned score function ŝ ≈ ∇ log p, we now consider discretizing (3.3). In this work we make use of the exponential integrator (EI), which fixes a sequence (to be specified) of reverse process timesteps 0 = τ N -1. (3.4) Recently, building off of the works by Chen et al. (2023a) and Lee et al. (2023), Benton et al. (2024) showed that it suffices to control the score approximation error in L 2 (p t ) to guarantee that the process (3.4) yields a high quality sample from p 0 . 2 Score function estimation. To estimate the score function ∇ log p t over the interval [0, T ], one would ideally minimize the least-squares objective over a model ŝ, R(ŝ) :=
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Much Ado About Noising: Dispelling the Myths of Generative Robotic ControlChaoyi Pan, Giridharan Anantharaman, Nai-Chieh Huang, Claire Jin et al.ICLR 2026 · 51 citations
- A solvable model of learning generative diffusion: theory and insightsHugo Cui, Cengiz Pehlevan, Yue M. LuNeurIPS 2025 · 11 citations
- Dimension-free Score Matching and Time Bootstrapping for Diffusion ModelsSyamantak Kumar, Dheeraj Nagaraj, Purnamrita SarkarNeurIPS 2025 · 2 citations
- Diffusion Models Are Statistically Optimal for Learning Low-Dimensional Multi-Modal DistributionsJingda Wu, Changxiao CaiICML 2026 · 1 citation
- On the Interpolation Effect of Score Smoothing in Diffusion ModelsZhengdao ChenICLR 2026
Builds on28
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 3,959 citations
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang et al.NeurIPS 2022 · 1,546 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
Related papers
- Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional DataMinshuo Chen, Kaixuan Huang, Tuo Zhao, Mengdi WangICML 2023 · 168 citations
- Fast Sampling of Diffusion Models with Exponential IntegratorQinsheng Zhang, Yongxin ChenICLR 2023 · 58 citations
- On the Generalization Properties of Diffusion ModelsPuheng Li, Zhong Li, Huishuai Zhang, Jiang BianNeurIPS 2023 · 86 citations
- O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal AssumptionsGen Li, Yuling YanICLR 2025 · 1 citation
- High-accuracy and dimension-free sampling with diffusionsKhashayar Gatmiry, Sitan Chen, Adil SalimICML 2026 · 7 citations
