Dynamic Dual-Output Diffusion Models
Yaniv Benny, Lior Wolf
摘要
Iterative denoising-based generation, also known as denoising diffusion models, has recently been shown to be comparable in quality to other classes of generative models, and even surpass them. Including, in particular, Generative Adversarial Networks, which are currently the state of the art in many sub-tasks of image generation. However, a major drawback of this method is that it requires hundreds of iterations to produce a competitive result. Recent works have proposed solutions that allow for faster generation with fewer iterations, but the image quality gradually deteriorates with increasingly fewer iterations being applied during generation. In this paper, we reveal some of the causes that affect the generation quality of diffusion models, especially when sampling with few iterations, and come up with a simple, yet effective, solution to mitigate them. We consider two opposite equations for the iterative denoising, the first predicts the applied noise, and the second predicts the image directly. Our solution takes the two options and learns to dynamically alternate between them through the denoising process. Our proposed solution is general and can be applied to any existing diffusion model. As we show, when applied to various SOTA architectures, our solution immediately improves their generation quality, with negligible added complexity and parameters. We experiment on multiple datasets and configurations and run an extensive ablation study to support these findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Non-autoregressive Conditional Diffusion Models for Time Series PredictionLifeng Shen, James T. KwokICML 2023 · 被引用 128 次
- Elucidating the Exposure Bias in Diffusion ModelsMang Ning, Mingxiao Li, Jianlin Su, Albert Ali Salah 等ICLR 2024 · 被引用 95 次
- Multi-Resolution Diffusion Models for Time Series ForecastingLifeng Shen, Weiyu Chen, James T. KwokICLR 2024 · 被引用 54 次
- Efficient Conditional Diffusion Model with Probability Flow Sampling for Image Super-resolutionYutao Yuan, Chun YuanAAAI 2024 · 被引用 16 次
- Latent-to-Data Cascaded Diffusion Models for Unconditional Time Series GenerationLifeng Shen, Kai Syun Hou, Weiyu Chen, James T. KwokICLR 2026 · 被引用 15 次
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
相关 Paper
- Truncated Diffusion Probabilistic Models and Diffusion-based Adversarial Auto-EncodersHuangjie Zheng, Pengcheng He, Weizhu Chen, Mingyuan ZhouICLR 2023 · 被引用 19 次
- One Step Diffusion via Shortcut ModelsKevin Frans, Danijar Hafner, Sergey Levine, Pieter AbbeelICLR 2025 · 被引用 2 次
- Fast constrained sampling in pre-trained diffusion modelsAlexandros Graikos, Nebojsa Jojic, Dimitris SamarasNeurIPS 2025 · 被引用 9 次
- OMS-DPM: Optimizing the Model Schedule for Diffusion Probabilistic ModelsEnshu Liu, Xuefei Ning, Zinan Lin, Huazhong Yang 等ICML 2023 · 被引用 57 次
- Factorized Diffusion Architectures for Unsupervised Image Generation and SegmentationXin Yuan, Michael MaireNeurIPS 2024 · 被引用 4 次
