Reverse Transition Kernel: A Flexible Framework to Accelerate Diffusion Inference
Xunpeng Huang, Difan Zou, Hanze Dong, Yi Zhang, Yian Ma, Tong Zhang
摘要
To generate data from trained diffusion models, most inference algorithms, such as DDPM, DDIM, and other variants, rely on discretizing the reverse SDEs or their equivalent ODEs. In this paper, we view such approaches as decomposing the entire denoising diffusion process into several segments, each corresponding to a reverse transition kernel (RTK) sampling subproblem. Specifically, DDPM uses a Gaussian approximation for the RTK, resulting in low per-subproblem complexity but requiring a large number of segments (i.e., subproblems), which is conjectured to be inefficient. To address this, we develop a general RTK framework that enables a more balanced subproblem decomposition, resulting in subproblems, each with strongly log-concave targets. We then propose leveraging two fast sampling algorithms, the Metropolis-Adjusted Langevin Algorithm (MALA) and Underdamped Langevin Dynamics (ULD), for solving these strongly log-concave subproblems. This gives rise to the RTK-MALA and RTK-ULD algorithms for diffusion inference. In theory, we further develop the convergence guarantees for RTK-MALA and RTK-ULD in total variation (TV) distance: RTK-ULD can achieve target error within under mild conditions, and RTK-MALA enjoys a convergence rate under slightly stricter conditions. These theoretical results surpass the state-of-the-art convergence rates for diffusion inference and are well supported by numerical experiments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Dimension-free convergence of diffusion models for approximate Gaussian mixturesGen Li, Changxiao Cai, Yuting WeiICML 2026 · 被引用 20 次
- High-accuracy sampling for diffusion models and log-concave distributionsFan Chen, Sinho Chewi, Constantinos Daskalakis, Alexander RakhlinICML 2026 · 被引用 12 次
- Machine Unlearning in 3D Generation: A Perspective-Coherent Acceleration FrameworkShixuan Wang, Jingwen Ye, Xinchao WangNeurIPS 2025 · 被引用 5 次
- Are First-Order Diffusion Samplers Really Slower? A Fast Forward-Value ApproachYuchen Jiao, Na Li, Changxiao Cai, Gen LiICML 2026 · 被引用 1 次
- Unified Convergence Analysis for Score-Based Diffusion Models with Deterministic SamplersRunjia Li, Qiwei Di, Quanquan GuICLR 2025
它引用的顶会 Paper17
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 StepsCheng Lu, Yuhao Zhou, Fan Bao, Jianfei Chen 等NeurIPS 2022 · 被引用 2,653 次
相关 Paper
- Faster high-accuracy log-concave sampling via algorithmic warm startsJason M. Altschuler, Sinho ChewiFOCS 2023 · 被引用 6 次
- O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal AssumptionsGen Li, Yuling YanICLR 2025 · 被引用 1 次
- Accelerating Convergence of Score-Based Diffusion Models, ProvablyGen Li, Yu Huang, Timofey Efimov, Yuting Wei 等ICML 2024 · 被引用 75 次
- The Poisson Midpoint Method for Langevin Dynamics: Provably Efficient Discretization for Diffusion ModelsSaravanan Kandasamy, Dheeraj NagarajNeurIPS 2024 · 被引用 14 次
- Faster Diffusion Sampling with Randomized Midpoints: Sequential and ParallelShivam Gupta, Linda Cai, Sitan ChenICLR 2025
