Neural Network-Based Score Estimation in Diffusion Models: Optimization and Generalization
Yinbin Han, Meisam Razaviyayn, Renyuan Xu
摘要
Diffusion models have emerged as a dominant paradigm in generative AI, rivaling GANs in producing high-fidelity and robust samples. A core component of these models is learning the score function of perturbed data distribution via denoising score matching. While recent theoretical works have established strong statistical guarantees for diffusion models, they predominantly rely on algorithm-agnostic assumptions, presuming access to a theoretical oracle that perfectly minimizes the empirical risk. In practice, however, score functions are parameterized by highly non-convex neural networks and trained via gradient descent (GD). It remains a major open question whether practical gradient-based algorithms can navigate the optimization landscape of score matching to achieve provable accuracy. As a first step toward answering this question, this paper establishes a mathematical framework for analyzing score estimation using neural networks trained by GD. Our analysis covers both the optimization and the generalization aspects of the learning procedure. In particular, we propose a novel parametric formulation that reduces denoising score matching to a regression problem with inherently noisy labels. Unlike standard supervised learning, the score-matching problem introduces unique theoretical challenges, including unbounded input, vector-valued output, and an additional time variable, preventing existing techniques from being applied directly. We address these challenges by showing that, with proper designs, the evolution of GD-trained neural networks can be accurately approximated by a sequence of localized kernel regression problems. Our analysis is grounded in a novel parametric form of the neural network and an innovative connection between score matching and regression analysis, which facilitate the application of advanced statistical and optimization techniques. Furthermore, since prolonged training on noisy labels causes catastrophic overfitting, we derive a novel extension of early-stopping rules for unbounded domains. This, in turn, allows us to establish the first minimax-optimal generalization error (sample complexity) bounds for GD-trained neural networks in diffusion models. Finally, we validate our theory-inspired optimization framework on a real-world Credit Default dataset, demonstrating that our principled approach achieves performance comparable to heavily tuned heuristic training schemes in generating high-fidelity financial tabular data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Evaluating the design space of diffusion-based generative modelsYuqing Wang, Ye He, Molei TaoNeurIPS 2024 · 被引用 30 次
- Constrained Diffusion Models via Dual TrainingShervin Khalafi, Dongsheng Ding, Alejandro RibeiroNeurIPS 2024 · 被引用 24 次
- Unraveling the Smoothness Properties of Diffusion Models: A Gaussian Mixture PerspectiveYingyu Liang, Zhizhou Sha, Zhenmei Shi, Zhao Song 等ICCV 2025 · 被引用 23 次
- A solvable model of learning generative diffusion: theory and insightsHugo Cui, Cengiz Pehlevan, Yue M. LuNeurIPS 2025 · 被引用 11 次
- Algorithm- and Data-Dependent Generalization Bounds for Diffusion ModelsBenjamin Dupuis, Dario Shariatian, Maxime Haddouche, Alain Durmus 等NeurIPS 2025 · 被引用 5 次
它引用的顶会 Paper29
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Improved Techniques for Training Score-Based Generative ModelsYang Song, Stefano ErmonNeurIPS 2020 · 被引用 1,527 次
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung 等ICLR 2020 · 被引用 1,166 次
- Score-based Generative Modeling in Latent SpaceArash Vahdat, Karsten Kreis, Jan KautzNeurIPS 2021 · 被引用 903 次
相关 Paper
- Implicit Regularisation in Diffusion Models: An Algorithm-Dependent Generalisation AnalysisTyler Farghly, Patrick Rebeschini, George Deligiannidis, Arnaud DoucetICLR 2026 · 被引用 9 次
- Convergence Dynamics of Over-Parameterized Score Matching for a Single GaussianYiran Zhang, Weihang Xu, Mo Zhou, Maryam Fazel 等ICLR 2026 · 被引用 2 次
- Diffusion Based Representation LearningSarthak Mittal, Korbinian Abstreiter, Stefan Bauer, Bernhard Schölkopf 等ICML 2023 · 被引用 71 次
- Understanding Generalizability of Diffusion Models Requires Rethinking the Hidden Gaussian StructureXiang Li, Yixiang Dai, Qing QuNeurIPS 2024 · 被引用 45 次
- Maximum Likelihood Training for Score-based Diffusion ODEs by High Order Denoising Score MatchingCheng Lu, Kaiwen Zheng, Fan Bao, Jianfei Chen 等ICML 2022 · 被引用 109 次
