What Exactly Does Guidance Do in Masked Discrete Diffusion Models
Ye He, Kevin Rojas, Molei Tao
摘要
Masked discrete diffusion models have been gaining popularity recently, and classifier-free guidance, just like its continuous counterpart, has been proposed to enable efficacious conditional generation by discrete diffusion. To quantify the precise effect of discrete guidance, this article considers masked discrete diffusion with arbitrary data distribution in low dimension, so that the distribution that guided masked discrete diffusion samples from, as well as the sampling dynamics, can be analytically and exactly quantified and interpreted. When the full data distribution is a mixture over classes and the goal is to sample from a specific class, guidance amplifies class-specific regions while suppresses regions shared with other classes. This effect depends on the guidance strength and induces distinct covariance structures in the sampled distribution. Notably, we observe quantitatively different behaviors in D and D. We also show that for large , the decay rate of the total variation () along the reverse dynamics is double-exponential in for both D and D. These findings highlight the role of guidance, not just in shaping the output distribution, but also in controlling the dynamics of the sampling trajectory. Our theoretical analysis is supported by experiments that illustrate the geometric effects of guidance and its impact on convergence.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- MDNS: Masked Diffusion Neural Sampler via Stochastic Optimal ControlYuchen Zhu, Wei Guo, Jaemoo Choi, Guan-Horng Liu 等NeurIPS 2025 · 被引用 24 次
- Diffuse Everything: Multimodal Diffusion Models on Arbitrary State SpacesKevin Rojas, Yuchen Zhu, Sichen Zhu, Felix X.-F. Ye 等ICML 2025
它引用的顶会 Paper25
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
相关 Paper
- Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim ImpactKevin Rojas, Ye He, Chieh-Hsin Lai, Yuhta Takida 等ICLR 2026 · 被引用 11 次
- Stage-wise Dynamics of Classifier-Free Guidance in Diffusion ModelsCheng Jin, Qitan Shi, Yuantao GuICLR 2026 · 被引用 13 次
- Provable Efficiency of Guidance in Diffusion Models for General Data DistributionGen Li, Yuchen JiaoICML 2025
- Overshoot and Shrinkage in Classifier-Free Guidance: From Theory to PracticeKrunoslav Lehman Pavasovic, Jakob Verbeek, Giulio Biroli, Marc MezardICLR 2026
- Inner Classifier-Free Guidance and Its Taylor Expansion for Diffusion ModelsShikun Sun, Longhui Wei, Zhicai Wang, Zixuan Wang 等ICLR 2024 · 被引用 2 次
