Supercharged One-Step Text-to-Image Diffusion Models with Negative Prompts
Viet Nguyen, Anh Nguyen, Trung Dao, Khoi Nguyen, Cuong Pham, Toan Tran, Anh Tran
摘要
The escalating demand for real-time image synthesis has driven significant advancements in one-step diffusion models, which inherently offer expedited generation speeds compared to traditional multi-step methods. However, this enhanced efficiency is frequently accompanied by a compromise in the controllability of image attributes. While negative prompting, typically implemented via classifierfree guidance (CFG), has proven effective for fine-grained control in multi-step models, its application to one-step generators remains largely unaddressed. Due to the lack of iterative refinement, as in multi-step diffusion, directly applying CFG to one-step generation leads to blending artifacts and diminished output quality. To fill this gap, we introduce Negative-Away Steer Attention (NASA), an efficient method that integrates negative prompts into one-step diffusion models. NASA operates within the intermediate representation space by leveraging cross-attention mechanisms to suppress undesired visual attributes. This strategy avoids the blending artifacts inherent in output-space guidance and achieves high efficiency, incurring only a minimal 1.89% increase in FLOPs compared to the computational doubling of CFG. Furthermore, NASA can be seamlessly integrated into existing timestep distillation frameworks, enhancing the student's output quality. Experimental results demonstrate that NASA substantially improves controllability and output quality, achieving an HPSv2 score of 31.21, setting a new state-of-the-art benchmark for one-step diffusion models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Normalized Attention Guidance: Universal Negative Guidance for Diffusion ModelsDar-Yen Chen, Hmrishav Bandyopadhyay, Kai Zou, Yi-Zhe SongNeurIPS 2025 · 被引用 17 次
- Anti-I2V: Safeguarding your Photos from Malicious Image-to-video GenerationDuc Vu, Anh Nguyen, Chi Tran, Anh TranCVPR 2026 · 被引用 7 次
- InverFill: One-Step Inversion for Enhanced Few-Step Diffusion InpaintingDuc Vu, Kien Nguyen, Trong-Tung Nguyen, Ngan Nguyen 等CVPR 2026 · 被引用 4 次
- Restoring Initial Noise Sensitivity in Text-to-Image Distillation through Geometric Alignmenthuayang Huang, Ruoyu Wang, Jinhui Zhao, Wei Deng 等ICML 2026
它引用的顶会 Paper27
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
相关 Paper
- VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign FlipWenqi Guo, Shan DuICLR 2026 · 被引用 3 次
- Adaptive Guidance: Training-free Acceleration of Conditional Diffusion ModelsAngela Castillo, Jonas Kohler, Juan C. Pérez, Juan Pablo Pérez 等AAAI 2025 · 被引用 1 次
- Adding Additional Control to One-Step Diffusion with Joint Distribution MatchingYihong Luo, Tianyang Hu, Yifan Song, Jiacheng Sun 等ICCV 2025
- Stylekeeper: Prevent Content Leakage using Negative Visual Query GuidanceJaeseok Jeong, Junho Kim, Gayoung Lee, Yunjey Choi 等ICCV 2025
- Realism Control One-step Diffusion for Real-world Image Super ResolutionZongliang Wu, Siming Zheng, Peng-Tao Jiang, Xin YuanAAAI 2026
