Towards Understanding the Mechanisms of Classifier-Free Guidance
Xiang Li, Rongrong Wang, Qing Qu
摘要
systems, yet its underlying mechanisms remain poorly understood. In this work, we begin by analyzing CFG in a simplified linear diffusion model, where we show its behavior closely resembles that observed in the nonlinear case. Our analysis reveals that linear CFG improves generation quality via three distinct components: (i) a mean-shift term that approximately steers samples in the direction of class means, (ii) a positive Contrastive Principal Components (CPC) term that amplifies class-specific features, and (iii) a negative CPC term that suppresses generic features prevalent in unconditional data. We then verify these insights in real-world, nonlinear diffusion models: over a broad range of noise levels, linear CFG resembles the behavior of its nonlinear counterpart. Although the two eventually diverge at low noise levels, we discuss how the insights from the linear analysis still shed light on the CFG's mechanism in the nonlinear regime.
Contributions. Our main contributions are as follows:
• We identify the lack of class-specificity issue of naive conditional sampling, linking it to the nondistinctiveness of class covariances. Under a linear model assumption, we show CFG overcomes this issue by amplifying class-specific features, suppressing unconditional ones and shifting the samples in the direction of class mean.
• We validate these insights derived in the linear model on real diffusion models, demonstrating that:
(i) at high to moderate noise levels, linear CFG closely matches the effects of nonlinear CFG, and (ii) at low noise levels, the insights from the linear analysis can still shed light on the mechanism of CFG in this nonlinear regime.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- A Closer Look at Model Collapse: From a Generalization-to-Memorization PerspectiveLianghe Shi, Meng Wu, Huijie Zhang, Zekai Zhang 等NeurIPS 2025 · 被引用 22 次
- Generalization of Diffusion Models Arises with a Balanced Representation SpaceZekai Zhang, Xiao Li, Xiang Li, Lianghe Shi 等ICLR 2026 · 被引用 14 次
- Stage-wise Dynamics of Classifier-Free Guidance in Diffusion ModelsCheng Jin, Qitan Shi, Yuantao GuICLR 2026 · 被引用 13 次
- General and Efficient Steering of Unconditional Diffusion ModelsQingsong Wang, Misha Belkin, Yusu WangICML 2026
它引用的顶会 Paper22
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 被引用 3,959 次
相关 Paper
- Overshoot and Shrinkage in Classifier-Free Guidance: From Theory to PracticeKrunoslav Lehman Pavasovic, Jakob Verbeek, Giulio Biroli, Marc MezardICLR 2026
- Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim ImpactKevin Rojas, Ye He, Chieh-Hsin Lai, Yuhta Takida 等ICLR 2026 · 被引用 11 次
- ContrastiveCFG: Guiding Diffusion Sampling by Contrasting Positive and Negative ConceptsJinho Chang, Changsun Lee, Hyungjin Chung, Jong Chul YEICML 2026
- Provable Efficiency of Guidance in Diffusion Models for General Data DistributionGen Li, Yuchen JiaoICML 2025
- No Training, No Problem: Rethinking Classifier-Free Guidance for Diffusion ModelsSeyedmorteza Sadat, Manuel Kansy, Otmar Hilliges, Romann M. WeberICLR 2025
