Soft Truncation: A Universal Training Technique of Score-based Diffusion Model for High Precision Score Estimation
Dongjun Kim, Seungjae Shin, Kyungwoo Song, Wanmo Kang, Il-Chul Moon
Abstract
Recent advances in diffusion models bring state-of-the-art performance on image generation tasks. However, empirical results from previous research in diffusion models imply an inverse correlation between density estimation and sample generation performances. This paper investigates with sufficient empirical evidence that such inverse correlation happens because density estimation is significantly contributed by small diffusion time, whereas sample generation mainly depends on large diffusion time. However, training a score network well across the entire diffusion time is demanding because the loss scale is significantly imbalanced at each diffusion time. For successful training, therefore, we introduce Soft Truncation, a universally applicable training technique for diffusion models, that softens the fixed and static truncation hyperparameter into a random variable. In experiments, Soft Truncation achieves state-of-the-art performance on CIFAR-10, CelebA, CelebA-HQ 256x256, and STL-10 datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d9bef5aa-5dad-459d-8657-0841607fd3c7Cited by top-tier papers61
- Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of DiffusionDongjun Kim, Chieh-Hsin Lai, Wei-Hsiang Liao, Naoki Murata et al.ICLR 2024 · 377 citations
- Masked Diffusion Transformer is a Strong Image SynthesizerShanghua Gao, Pan Zhou, Ming-Ming Cheng, Shuicheng YanICCV 2023 · 290 citations
- Patch Diffusion: Faster and More Data-Efficient Training of Diffusion ModelsZhendong Wang, Yifan Jiang, Huangjie Zheng, Peihao Wang et al.NeurIPS 2023 · 205 citations
- Solving Linear Inverse Problems Provably via Posterior Sampling with Latent Diffusion ModelsLitu Rout, Negin Raoof, Giannis Daras, Constantine Caramanis et al.NeurIPS 2023 · 193 citations
- Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional DataMinshuo Chen, Kaixuan Huang, Tuo Zhao, Mengdi WangICML 2023 · 168 citations
Builds on15
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Improved Techniques for Training Score-Based Generative ModelsYang Song, Stefano ErmonNeurIPS 2020 · 1,527 citations
Related papers
- A Simple Early Exiting Framework for Accelerated Sampling in Diffusion ModelsTae Hong Moon, Moonseok Choi, EungGu Yun, Jongmin Yoon et al.ICML 2024 · 10 citations
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 3,959 citations
- Training Unbiased Diffusion Models From Biased DatasetYeongmin Kim, Byeonghu Na, Minsang Park, JoonHo Jang et al.ICLR 2024 · 37 citations
- Rényi Diffusion ModelsYirong Shen, Lu GAN, Cong LingICML 2026 · 4 citations
- Maximum Likelihood Training for Score-based Diffusion ODEs by High Order Denoising Score MatchingCheng Lu, Kaiwen Zheng, Fan Bao, Jianfei Chen et al.ICML 2022 · 109 citations
