Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models
Dar-Yen Chen, Hmrishav Bandyopadhyay, Kai Zou, Yi-Zhe Song
Abstract
Negative guidance -explicitly suppressing unwanted attributes -remains a fundamental challenge in diffusion models, particularly in few-step sampling regimes. While Classifier-Free Guidance (CFG) works well in standard settings, it fails under aggressive sampling step compression due to divergent predictions between positive and negative branches. We present Normalized Attention Guidance (NAG), an efficient, training-free mechanism that applies extrapolation in attention space with L1-based normalization and refinement. NAG restores effective negative guidance where CFG collapses while maintaining fidelity. Unlike existing approaches, NAG generalizes across architectures (UNet, DiT), sampling regimes (few-step, multi-step), and modalities (image, video), functioning as a universal plug-in with minimal computational overhead. Through extensive experimentation, we demonstrate consistent improvements in text alignment (CLIP Score), fidelity (FID, PFID), and human-perceived quality (ImageReward). Our ablation studies validate each design component, while user studies confirm significant preference for NAG-guided outputs. As a model-agnostic inference-time approach requiring no retraining, NAG provides effortless negative guidance for all modern diffusion frameworks -pseudocode in the Appendix! -Cow A tiger cow -Tiger Sketch of UFO over pyramid -Realistic, complex -Black and white -Blurry, low contrast Photo of aurora -Green -Male Portrait of AI researcher -Glasses Negative Prompting Figure 1: Negative prompting on 4-step Flux-Schnell [1]. CFG fails in few-step models. NAG restores effective negative prompting, enabling direct suppression of visual, semantic, and stylistic attributes, such as "glasses," "tiger," "realistic," or "blurry." This enhances controllability and expands creative freedom across composition, style, and quality-including prompt-based debiasing. 39th Conference on Neural Information Processing Systems (NeurIPS 2025).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 38a48d0a-dfd2-4854-849e-ab52335f62d0Cited by top-tier papers6
- Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion TransformersRuidong Chen, Yancheng Bai, Xuanpu Zhang, Jianhao Zeng et al.CVPR 2026 · 9 citations
- VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign FlipWenqi Guo, Shan DuICLR 2026 · 3 citations
- Protosampling: Enabling Free-Form Convergence of Sampling and Prototyping through Canvas-Driven Visual AI GenerationAlicia Guo, David Ledo, George W. Fitzmaurice, Fraser AndersonCHI 2026 · 2 citations
- GuidedBridge: Training-freely Improving Bridge Models with Prior GuidanceZehua Chen, Yucheng Yang, Binjie Yuan, Kaiwen Zheng et al.ICML 2026
- Rethinking Global Text Conditioning in Diffusion TransformersNikita Starodubcev, Daniil Pakhomov, Zongze Wu, Ilya Drobyshevskiy et al.ICLR 2026
Builds on39
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- Supercharged One-Step Text-to-Image Diffusion Models with Negative PromptsViet Nguyen, Anh Nguyen, Trung Dao, Khoi Nguyen et al.ICCV 2025 · 1 citation
- Guiding Diffusion Models with Semantically Degraded ConditionsShilong Han, Yuming Zhang, Hongxia WangCVPR 2026 · 1 citation
- Diffusion-NPO: Negative Preference Optimization for Better Preference Aligned Generation of Diffusion ModelsFu-Yun Wang, Yunhao Shui, Jingtan Piao, Keqiang Sun et al.ICLR 2025
- Guiding Diffusion Models With Adaptive Negative Sampling Without External ResourcesAlakh Desai, Nuno VasconcelosICCV 2025 · 1 citation
- Adaptive Guidance: Training-free Acceleration of Conditional Diffusion ModelsAngela Castillo, Jonas Kohler, Juan C. Pérez, Juan Pablo Pérez et al.AAAI 2025 · 1 citation
