ShiftDDPMs: Exploring Conditional Diffusion Models by Shifting Diffusion Trajectories
Zijian Zhang, Zhou Zhao, Jun Yu, Qi Tian
Abstract
Diffusion models have recently exhibited remarkable abilities to synthesize striking image samples since the introduction of denoising diffusion probabilistic models (DDPMs). Their key idea is to disrupt images into noise through a fixed forward process and learn its reverse process to generate samples from noise in a denoising way. For conditional DDPMs, most existing practices relate conditions only to the reverse process and fit it to the reversal of unconditional forward process. We find this will limit the condition modeling and generation in a small time window. In this paper, we propose a novel and flexible conditional diffusion model by introducing conditions into the forward process. We utilize extra latent space to allocate an exclusive diffusion trajectory for each condition based on some shifting rules, which will disperse condition modeling to all timesteps and improve the learning capacity of model. We formulate our method, which we call ShiftDDPMs, and provide a unified point of view on existing related methods. Extensive qualitative and quantitative experiments on image synthesis demonstrate the feasibility and effectiveness of ShiftDDPMs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b837011c-38cf-420a-a2d0-5eb4018dac2bCited by top-tier papers7
- Neural Flow Diffusion Models: Learnable Forward Process for Improved Diffusion ModellingGrigory Bartosh, Dmitry P. Vetrov, Christian Andersson NaessethNeurIPS 2024 · 49 citations
- Long-tailed Diffusion Models with Oriented CalibrationTianjiao Zhang, Huangjie Zheng, Jiangchao Yao, Xiangfeng Wang et al.ICLR 2024 · 22 citations
- Characteristic Guidance: Non-linear Correction for Diffusion Model at Large Guidance ScaleCandi Zheng, Yuan LanICML 2024 · 18 citations
- TranSpeech: Speech-to-Speech Translation With Bilateral PerturbationRongjie Huang, Jinglin Liu, Huadai Liu, Yi Ren et al.ICLR 2023 · 17 citations
- Terrain Diffusion Network: Climatic-Aware Terrain Generation with Geological Sketch GuidanceZexin Hu, Kun Hu, Clinton Mo, Lei Pan et al.AAAI 2024 · 8 citations
Builds on13
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
Related papers
- VideoFusion: Decomposed Diffusion Models for High-Quality Video GenerationCVPR 2023
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- A Unified Conditional Framework for Diffusion-based Image RestorationYi Zhang, Xiaoyu Shi, Dasong Li, Xiaogang Wang et al.NeurIPS 2023 · 44 citations
- Going beyond Compositions, DDPMs Can Produce Zero-Shot InterpolationsJustin Deschenaux, Igor Krawczuk, Grigorios Chrysos, Volkan CevherICML 2024 · 6 citations
- Restoration based Generative ModelsJaemoo Choi, Yesom Park, Myungjoo KangICML 2023 · 5 citations
