Structure-Guided Adversarial Training of Diffusion Models
Ling Yang, Haotian Qian, Zhilong Zhang, Jingwei Liu, Bin Cui
Abstract
Diffusion models have demonstrated exceptional efficacy in various generative applications. While existing models focus on minimizing a weighted sum of denoising score matching losses for data distribution modeling, their training primarily emphasizes instance-level optimization, overlooking valuable structural information within each minibatch, indicative of pair-wise relationships among samples. To address this limitation, we introduce Structure-guided Adversarial training of Diffusion Models (SADM). In this pioneering approach, we compel the model to learn manifold structures between samples in each training batch. To ensure the model captures authentic manifold structures in the data distribution, we advocate adversarial training of the diffusion generator against a novel structure discriminator in a minimax game, distinguishing real manifold structures from the generated ones. SADM substantially outperforms existing methods in image generation and cross-domain fine-tuning tasks across 12 datasets, establishing a new state-of-the-art FID of 1.58 and 2.11 on ImageNet for classconditional image generation at resolutions of 256×256 and 512×512, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMsLing Yang, Zhaochen Yu, Chenlin Meng, Minkai Xu et al.ICML 2024 · 231 citations
- Cross-Modal Contextualized Diffusion Models for Text-Guided Visual Generation and EditingLing Yang, Zhilong Zhang, Zhaochen Yu, Jingwei Liu et al.ICLR 2024 · 25 citations
- Generative Adversarial DiffusionU-Chae Jun, Jaeeun Ko, Jiwoo KangICCV 2025 · 2 citations
- Mitigating Error Amplification in Fast Adversarial TrainingMengnan Zhao, Lihe Zhang, Bo Wang, Tianhang Zheng et al.CVPR 2026 · 1 citation
- Why Adversarially Train Diffusion Models?Maria Rosaria Briglia, Mujtaba Hussain Mirza, Giuseppe Lisanti, Iacopo MasiICLR 2026
Builds on40
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- SDDM: Score-Decomposed Diffusion Models on Manifolds for Unpaired Image-to-Image TranslationShikun Sun, Longhui Wei, Junliang Xing, Jia Jia et al.ICML 2023 · 21 citations
- Adversarial Score identity Distillation: Rapidly Surpassing the Teacher in One StepMingyuan Zhou, Huangjie Zheng, Yi Gu, Zhendong Wang et al.ICLR 2025
- Refining Generative Process with Discriminator Guidance in Score-based Diffusion ModelsDongjun Kim, Yeongmin Kim, Se Jung Kwon, Wanmo Kang et al.ICML 2023 · 109 citations
- Contrastive Flow MatchingGeorge Stoica, Vivek Ramanujan, Xiang Fan, Ali Farhadi et al.ICCV 2025 · 6 citations
- Analyzing and Improving the Training Dynamics of Diffusion ModelsTero Karras, Miika Aittala, Jaakko Lehtinen, Janne Hellsten et al.CVPR 2024
