Visual Generation Without Guidance
Huayu Chen, Kai Jiang, Kaiwen Zheng, Jianfei Chen, Hang Su, Jun Zhu
Abstract
Classifier-Free Guidance (CFG) has been a default technique in various visual generative models, yet it requires inference from both conditional and unconditional models during sampling. We propose to build visual models that are free from guided sampling. The resulting algorithm, Guidance-Free Training (GFT), matches the performance of CFG while reducing sampling to a single model, halving the computational cost. Unlike previous distillation-based approaches that rely on pretrained CFG networks, GFT enables training directly from scratch. GFT is simple to implement. It retains the same maximum likelihood objective as CFG and differs mainly in the parameterization of conditional models. Implementing GFT requires only minimal modifications to existing codebases, as most design choices and hyperparameters are directly inherited from CFG. Our extensive experiments across five distinct visual models demonstrate the effectiveness and versatility of GFT. Across domains of diffusion, autoregressive, and maskedprediction modeling, GFT consistently achieves comparable or even lower FID scores, with similar diversity-fidelity trade-offs compared with CFG baselines, all while being guidance-free. Code: https://github.com/thu-ml/GFT .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Improved Mean Flows: On the Challenges of Fastforward Generative ModelsZhengyang Geng, Yiyang Lu, Zongze Wu, Eli Shechtman et al.CVPR 2026 · 116 citations
- Terminal Velocity MatchingLinqi Zhou, Mathias Parger, Ayaan Haque, Jiaming SongICLR 2026 · 15 citations
- Bidirectional Normalizing Flow: From Data to Noise and BackYiyang Lu, Qiao Sun, Xianbang Wang, Zhicheng Jiang et al.CVPR 2026 · 7 citations
- DogFit: Domain-guided Fine-tuning for Efficient Transfer Learning of Diffusion ModelsYara Bahram, Mohammadhadi Shateri, Eric GrangerAAAI 2026 · 4 citations
- Improving Diffusion Generalization with Weak-to-Strong Segmented GuidanceLiangyu Yuan, Yufei Huang, Mingkun Lei, Tong Zhao et al.CVPR 2026 · 1 citation
Builds on44
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- No Training, No Problem: Rethinking Classifier-Free Guidance for Diffusion ModelsSeyedmorteza Sadat, Manuel Kansy, Otmar Hilliges, Romann M. WeberICLR 2025
- Toward Guidance-Free AR Visual Generation via Condition Contrastive AlignmentHuayu Chen, Hang Su, Peize Sun, Jun ZhuICLR 2025
- Plug-and-Play Diffusion DistillationYi-Ting Hsiao, Siavash Khodadadeh, Kevin Duarte, Wei-An Lin et al.CVPR 2024
- Adaptive Guidance: Training-free Acceleration of Conditional Diffusion ModelsAngela Castillo, Jonas Kohler, Juan C. Pérez, Juan Pablo Pérez et al.AAAI 2025 · 1 citation
- Inner Classifier-Free Guidance and Its Taylor Expansion for Diffusion ModelsShikun Sun, Longhui Wei, Zhicai Wang, Zixuan Wang et al.ICLR 2024 · 2 citations
