LIFT and PLACE: A Simple, Stable, and Effective Knowledge Distillation Framework for Lightweight Diffusion Models
Hyunsoo Han, Sangyeop Yeo, Jaejun Yoo
摘要
We demonstrate that in knowledge distillation for diffusion models, the teacher network's highly complex denoising process-stemming from its substantially larger capacity-poses a significant challenge for the student model to faithfully mimic. To address this problem, we propose a coarse-to-fine distillation framework with LInear FiTtingbased distillation (LIFT) and Piecewise Local Adaptive Coefficient Estimation (PLACE). First, LIFT decomposes the objective into a "coarse" alignment and a "fine" refinement. The student is then trained on coarse alignment before proceeding to hard refinement. Second, PLACE extends LIFT to address spatially non-uniform errors by partitioning outputs into error-based groups, providing locally adaptive guidance. Our experiments show that LIFT and PLACE is effective across diffusion spaces (image/latent), backbones (U-Net/DiT), tasks (unconditional/conditional), datasets, and even extends to flow-based models such as MMDiT (SD3). Furthermore, under extreme compression with a 1.3M-parameter student (only 1.6% of the teacher), conventional KD fails to provide sufficient guidance for stable training, with FID scores often degrading to 50-200+, but our method remains stably convergent and achieves an FID of 15.73. Our project page is available at here.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper25
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari 等ICML 2024 · 被引用 3,620 次
- Consistency ModelsYang Song, Prafulla Dhariwal, Mark Chen, Ilya SutskeverICML 2023 · 被引用 1,720 次
相关 Paper
- Knowledge Diffusion for DistillationTao Huang, Yuan Zhang, Mingkai Zheng, Shan You 等NeurIPS 2023 · 被引用 125 次
- Reducing Spatial Fitting Error in Distillation of Denoising Diffusion ModelsShengzhe Zhou, Zejian Li, Shengyuan Zhang, Lefan Hou 等AAAI 2024 · 被引用 1 次
- Plug-and-Play Diffusion DistillationYi-Ting Hsiao, Siavash Khodadadeh, Kevin Duarte, Wei-An Lin 等CVPR 2024
- Flash Diffusion: Accelerating Any Conditional Diffusion Model for Few Steps Image GenerationClément Chadebec, Onur Tasar, Eyal Benaroche, Benjamin AubinAAAI 2025 · 被引用 52 次
- SenseFlow: Scaling Distribution Matching for Flow-based Text-to-Image DistillationXingtong Ge, Xin Zhang, Tongda Xu, Yi Zhang 等ICLR 2026 · 被引用 29 次
