Adaptive Non-Uniform Timestep Sampling for Accelerating Diffusion Model Training
Myunsoo Kim, Donghyeon Ki, Seong-Woong Shim, Byung-Jun Lee
摘要
As a highly expressive generative model, diffusion models have demonstrated exceptional success across various domains, including image generation, natural language processing, and combinatorial optimization. However, as data distributions grow more complex, training these models to convergence becomes increasingly computationally intensive. While diffusion models are typically trained using uniform timestep sampling, our research shows that the variance in stochastic gradients varies significantly across timesteps, with high-variance timesteps becoming bottlenecks that hinder faster convergence. To address this issue, we introduce a non-uniform timestep sampling method that prioritizes these more critical timesteps. Our method tracks the impact of gradient updates on the objective for each timestep, adaptively selecting those most likely to minimize the objective effectively. Experimental results demonstrate that this approach not only accelerates the training process, but also leads to improved performance at convergence. Furthermore, our method shows robust performance across various datasets, scheduling strategies, and diffusion architectures, outperforming previously proposed timestep sampling and weighting heuristics that lack this degree of robustness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Prior-Guided Diffusion Planning for Offline Reinforcement LearningDonghyeon Ki, JunHyeok Oh, Seong-Woong Shim, Byung-Jun LeeNeurIPS 2025 · 被引用 16 次
- T-LoRA: Single Image Diffusion Model Customization Without OverfittingVera Soboleva, Aibek Alanov, Andrey Kuznetsov, Konstantin SobolevAAAI 2026 · 被引用 9 次
- Fast3Dcache: Training-free 3D Geometry Synthesis AccelerationMengyu Yang, Yanming Yang, Chenyi Xu, Chenxi Song 等CVPR 2026 · 被引用 4 次
- Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image GenerationJohannes Schusterbauer, Ming Gui, Yusong Li, Pingchuan Ma 等CVPR 2026 · 被引用 4 次
- FALCON: False-Negative Aware Learning of Contrastive Negatives in Vision-Language AlignmentMyunsoo Kim, Seong-Woong Shim, Byung-Jun LeeCVPR 2026 · 被引用 2 次
它引用的顶会 Paper18
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 被引用 11,724 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 被引用 3,959 次
相关 Paper
- Non-uniform Timestep Sampling: Towards Faster Diffusion Model TrainingTianyi Zheng, Cong Geng, Peng-Tao Jiang, Ben Wan 等ACM MM 2024 · 被引用 7 次
- A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model TrainingKai Wang, Mingjia Shi, Yukun Zhou, Zekai Li 等CVPR 2025
- AutoDiffusion: Training-Free Optimization of Time Steps and Architectures for Automated Diffusion Model AccelerationLijiang Li, Huixia Li, Xiawu Zheng, Jie Wu 等ICCV 2023 · 被引用 83 次
- Align Your Steps: Optimizing Sampling Schedules in Diffusion ModelsAmirmojtaba Sabour, Sanja Fidler, Karsten KreisICML 2024 · 被引用 74 次
- Improved Noise Schedule for Diffusion TrainingTiankai Hang, Shuyang Gu, Jianmin Bao, Fangyun Wei 等ICCV 2025 · 被引用 5 次
