FaPS: A General and Fast Training Method for Diffusion Models
Xianglu Wang, Bangxian Han, Hu Ding
摘要
Diffusion models have achieved state-of-the-art performance in image generation tasks. However, training powerful diffusion models remains time-consuming, which limits their practical deployment. In this paper, we revisit the learning dynamics of diffusion models through the lens of spectral bias, a phenomenon in which deep neural networks prioritize learning low-frequency modes. Through an empirical analysis of diffusion training, we observe that diffusion models exhibit a dual spectral bias. First, over training iterations, they fit low-frequency components earlier than high-frequency details. Second, along the diffusion timesteps, early denoising steps mainly reconstruct coarse low-frequency content, while high-frequency details emerge in later steps. Motivated by this observation, we propose Frequency-aware Patch Selection (FaPS), a general and fast training method for diffusion models that can be applied to both UNet and DiT backbones. Specifically, FaPS introduces a frequency-aware gating that adaptively selects image patches based on their frequency information and focuses computation only on the selected patches. Since the selection decisions are discrete and thus non-differentiable, we model the gating as a stochastic policy network and optimize it end-to-end using a policy gradient method. Our experiments demonstrate that FaPS achieves up to faster training while maintaining comparable or superior generation quality, and improves the performance of diffusion models in limited-data settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper43
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
相关 Paper
- FreqTS: Frequency-Aware Token Selection for Accelerating Diffusion ModelsXinye Yang, Yuxin Yang, Haoran Pang, Aaron Xuxiang Tian 等AAAI 2025 · 被引用 2 次
- Frequency Regulation for Exposure Bias Mitigation in Diffusion ModelsMeng Yu, Kun ZhanACM MM 2025 · 被引用 1 次
- Elucidating the SNR-t Bias of Diffusion Probabilistic ModelsMeng Yu, Lei Sun, Jianhao Zeng, Xiangxiang Chu 等CVPR 2026 · 被引用 3 次
- Beyond Uniformity: Sample and Frequency Meta Weighting for Post-Training Quantization of Diffusion ModelsVan Cuong Pham, Anh Hoang, Cuong Nguyen, Trung Le 等ICLR 2026
- Diffusion Probabilistic Model Made SlimXingyi Yang, Daquan Zhou, Jiashi Feng, Xinchao WangCVPR 2023
