Lipschitz Singularities in Diffusion Models
Zhantao Yang, Ruili Feng, Han Zhang, Yujun Shen, Kai Zhu, Lianghua Huang, Yifei Zhang, Yu Liu, Deli Zhao, Jingren Zhou, Fan Cheng
摘要
Diffusion models, which employ stochastic differential equations to sample images through integrals, have emerged as a dominant class of generative models. However, the rationality of the diffusion process itself receives limited attention, leaving the question of whether the problem is well-posed and well-conditioned. In this paper, we explore a perplexing tendency of diffusion models: they often display the infinite Lipschitz property of the network with respect to time variable near the zero point. We provide theoretical proofs to illustrate the presence of infinite Lipschitz constants and empirical results to confirm it. The Lipschitz singularities pose a threat to the stability and accuracy during both the training and inference processes of diffusion models. Therefore, the mitigation of Lipschitz singularities holds great potential for enhancing the performance of diffusion models. To address this challenge, we propose a novel approach, dubbed E-TSDM, which alleviates the Lipschitz singularities of the diffusion model near the zero point of timesteps. Remarkably, our technique yields a substantial improvement in performance. Moreover, as a byproduct of our method, we achieve a dramatic reduction in the Fréchet Inception Distance of acceleration methods relying on network Lipschitz, including DDIM and DPM-Solver, by over 33%. Extensive experiments on diverse datasets validate our theory and method. Our work may advance the understanding of the general diffusion process, and also provide insights for the design of diffusion models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- DistillKac: Few-Step Image Generation via Damped Wave EquationsWeiqiao Han, Chenlin Meng, Christopher D Manning, Stefano ErmonICLR 2026 · 被引用 2 次
- Dataset Distillation as Pushforward Optimal QuantizationHong Ye Tan, Emma SladeICLR 2026 · 被引用 1 次
- Optimal Stopping in Latent Diffusion ModelsYu-Han Wu, Quentin Berthet, Gérard Biau, Claire Boyer 等ICML 2026 · 被引用 1 次
- Proof-of-Authorship for Diffusion-based AI Generated ContentDe Zhang Lee, Han Fang, Ee-Chien ChangCCS 2026
- SADA: Stability-guided Adaptive Diffusion AccelerationTing Jiang, Yixiao Wang, Hancheng Ye, Zishan Shao 等ICML 2025
它引用的顶会 Paper24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Towards More Accurate Diffusion Model Acceleration with a Timestep TunerMengfei Xia, Yujun Shen, Changsong Lei, Yu Zhou 等CVPR 2024 · 被引用 6 次
- Tackling the Singularities at the Endpoints of Time Intervals in Diffusion ModelsPengze Zhang, Hubery Yin, Chen Li, Xiaohua XieCVPR 2024 · 被引用 7 次
- Accelerating Convergence of Score-Based Diffusion Models, ProvablyGen Li, Yu Huang, Timofey Efimov, Yuting Wei 等ICML 2024 · 被引用 75 次
- Multi-Step Denoising Scheduled Sampling: Towards Alleviating Exposure Bias for Diffusion ModelsZhiyao Ren, Yibing Zhan, Liang Ding, Gaoang Wang 等AAAI 2024 · 被引用 15 次
- Schedule Your Edit: A Simple yet Effective Diffusion Noise Schedule for Image EditingHaonan Lin, Yan Chen, Jiahao Wang, Wenbin An 等NeurIPS 2024 · 被引用 46 次
