CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image Generation
Kangfu Mei, Mauricio Delbracio, Hossein Talebi, Zhengzhong Tu, Vishal M. Patel, Peyman Milanfar
摘要
Large generative diffusion models have revolution-ized text-to-image generation and offer immense po-tential for conditional generation tasks such as im-age enhancement, restoration, editing, and compositing. However, their widespread adoption is hindered by the high computational cost, which limits their real-time application. To address this challenge, we in-troduce a novel method dubbed CoDi, that adapts a pre-trained latent diffusion model to accept additional image conditioning inputs while significantly reducing the sampling steps required to achieve high-quality results. Our method can leverage architectures such as ControlNet to incorporate conditioning inputs with-out compromising the model's prior knowledge gained during large scale pre-training. Additionally, a con-ditional consistency loss enforces consistent predictions across diffusion steps, effectively compelling the model to generate high-quality images with conditions in a few steps. Our conditional-task learning and distil-lation approach outperforms previous distillation meth-ods, achieving a new state-of-the-art in producing high-quality images with very few steps (e.g., 1–4) across multiple tasks, including super-resolution, text-guided image editing, and depth-to-image generation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Learning Straight Flows: Variational Flow Matching for Efficient GenerationChenrui Ma, Xi Xiao, Tianyang Wang, Xiao Wang 等CVPR 2026 · 被引用 9 次
- Text-Aware Image Restoration with Diffusion ModelsJaewon Min, Jin Hyeon Kim, Paul Hyunbin Cho, Jaeeun Lee 等ICLR 2026 · 被引用 7 次
- Fine-Structure Preserved Real-World Image Super-Resolution Via Transfer Vae TrainingQiaosi Yi, Shuai Liu, Rongyuan Wu, Lingchen Sun 等ICCV 2025 · 被引用 4 次
- PatchScaler: An Efficient Patch-Independent Diffusion Model for Image Super-ResolutionYong Liu, Hang Dong, Jinshan Pan, Qingji Dong 等ICCV 2025 · 被引用 2 次
- Uni-DAD: Unified Distillation and Adaptation of Diffusion Models for Few-step Few-shot Image GenerationYara Bahram, Mélodie Desbos, Mohammadhadi Shateri, Eric GrangerCVPR 2026 · 被引用 1 次
它引用的顶会 Paper28
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Invertible Consistency Distillation for Text-Guided Image Editing in Around 7 StepsNikita Starodubcev, Mikhail Khoroshikh, Artem Babenko, Dmitry BaranchukNeurIPS 2024 · 被引用 18 次
- Realism Control One-step Diffusion for Real-world Image Super ResolutionZongliang Wu, Siming Zheng, Peng-Tao Jiang, Xin YuanAAAI 2026
- SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency DistillationJunsong Chen, Shuchen Xue, Yuyang Zhao, Jincheng Yu 等ICCV 2025 · 被引用 6 次
- Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion ModelHan Lin, Jaemin Cho, Abhay Zala, Mohit BansalICLR 2025
