Accelerating Diffusion Model Training under Minimal Budgets: A Condensation-Based Perspective
Rui Huang, Shitong Shao, Zikai Zhou, Pukun Zhao, Hangyu Guo, Tian Ye, Lichen Bai, Shuo Yang, Zeke Xie
摘要
Diffusion models have achieved remarkable performance on a wide range of generative tasks, yet training them from scratch is notoriously resource-intensive, typically requiring millions of training images and many GPU days. Motivated by a data-centric view of this bottleneck, we adopt a condensation-based perspective: given a large training set, the goal is to construct a much smaller condensed dataset that still supports training strong diffusion models under minimal data and compute budgets. To operationalize this perspective, we introduce Diffusion Dataset Condensation (D 2 C), a two-phase framework comprising Select and Attach. In the Select phase, a diffusion difficulty score combined with interval sampling is used to identify a compact, informative training subset from the original data. Building on this subset, the Attach phase further strengthens the conditional signals by augmenting each selected image with rich semantic and visual representations. To our knowledge, D 2 C is the first framework that systematically investigates dataset condensation for diffusion models, whereas prior condensation methods have mainly targeted discriminative architectures. Extensive experiments across data budgets (0.8%-8% of ImageNet), model architectures, and image resolutions demonstrate that D 2 C dramatically accelerates diffusion model training while preserving high generative quality. On ImageNet 256 2 with SiT-XL/2, D 2 C attains a FID of 4.3 in just 40k steps using only 0.8% of the training images, corresponding to about 233× and 100× faster training than vanilla SiT-XL/2 and SiT-XL/2 + REPA, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- UtilGen: Utility-Centric Generative Data Augmentation with Dual-Level Task AdaptationJiyu Guo, Shuo Yang, Yiming Huang, Yancheng Long 等NeurIPS 2025 · 被引用 4 次
- FastLightGen: Fast and Light Video Generation with Fewer Steps and ParametersShitong Shao, Yufei Gu, Zeke XieCVPR 2026 · 被引用 4 次
- LESA: Learnable Stage-Aware Predictors for Diffusion Model AccelerationPeiliang Cai, Jiacheng Liu, Haowen Xu, Xinyu Wang 等CVPR 2026 · 被引用 4 次
- CRAFT: Aligning Diffusion Models with Fine-Tuning Is Easier Than You ThinkZening Sun, Zhengpeng Xie, Lichen Bai, Shitong Shao 等CVPR 2026 · 被引用 3 次
- Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion ModelsZhirong Shen, Rui Huang, Jiacheng Liu, Chang Zou 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper38
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- FlashEval: Towards Fast and Accurate Evaluation of Text-to-Image Diffusion Generative ModelsLin Zhao, Tianchen Zhao, Zinan Lin, Xuefei Ning 等CVPR 2024 · 被引用 2 次
- Patch Diffusion: Faster and More Data-Efficient Training of Diffusion ModelsZhendong Wang, Yifan Jiang, Huangjie Zheng, Peihao Wang 等NeurIPS 2023 · 被引用 205 次
- Geometry-Aware Dataset Condensation for Diffusion Model TrainingXiao Cui, Yulei Qin, Mo Zhu, Wengang Zhou 等ICML 2026
- Diffusion Models and Semi-Supervised Learners Benefit Mutually with Few LabelsZebin You, Yong Zhong, Fan Bao, Jiacheng Sun 等NeurIPS 2023 · 被引用 61 次
- Progressive Distillation for Fast Sampling of Diffusion ModelsTim Salimans, Jonathan HoICLR 2022 · 被引用 9 次
