DEFT: Efficient Fine-tuning of Diffusion Models by Learning the Generalised -transform
Alexander Denker, Francisco Vargas, Shreyas Padhy, Kieran Didi, Simon V. Mathis, Riccardo Barbano, Vincent Dutordoir, Emile Mathieu, Urszula Julia Komorowska, Pietro Lió
摘要
Generative modelling paradigms based on denoising diffusion processes have emerged as a leading candidate for conditional sampling in inverse problems. In many real-world applications, we often have access to large, expensively trained unconditional diffusion models, which we aim to exploit for improving conditional sampling. Most recent approaches are motivated heuristically and lack a unifying framework, obscuring connections between them. Further, they often suffer from issues such as being very sensitive to hyperparameters, being expensive to train or needing access to weights hidden behind a closed API. In this work, we unify conditional training and sampling using the mathematically well-understood Doob's h-transform. This new perspective allows us to unify many existing methods under a common umbrella. Under this framework, we propose DEFT (Doob's h-transform Efficient FineTuning), a new approach for conditional generation that simply fine-tunes a very small network to quickly learn the conditional -transform, while keeping the larger unconditional network unchanged. DEFT is much faster than existing baselines while achieving state-of-the-art performance across a variety of linear and non-linear benchmarks. On image reconstruction tasks, we achieve speedups of up to 1.6, while having the best perceptual quality on natural images and reconstruction performance on medical images. Further, we also provide initial experiments on protein motif scaffolding and outperform reconstruction guidance methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Physics-Driven Spatiotemporal Modeling for AI-Generated Video DetectionShuhai Zhang, Zihao Lian, Jiahao Yang, Daiyuan Li 等NeurIPS 2025 · 被引用 29 次
- Meta Flow Maps enable scalable reward alignmentPeter Potaptchik, Adhi Saravanan, Abbas Mammadov, Alvaro Prat 等ICML 2026 · 被引用 25 次
- RNE: plug-and-play diffusion inference-time control and energy-based trainingJiajun He, José Miguel Hernández-Lobato, Yuanqi Du, Francisco VargasICLR 2026 · 被引用 17 次
- Value Gradient Guidance for Flow Matching AlignmentZhen Liu, Tim Z. Xiao, Carles Domingo-Enrich, Weiyang Liu 等NeurIPS 2025 · 被引用 15 次
- CREPE: Controlling diffusion with REPlica ExchangeJiajun He, Paul Jeha, Peter Potaptchik, Leo Zhang 等ICLR 2026 · 被引用 11 次
它引用的顶会 Paper39
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
相关 Paper
- Supervised Guidance Training for Infinite-Dimensional Diffusion ModelsElizabeth Baker, Alexander Denker, Jes FrellsenICML 2026 · 被引用 2 次
- h-Edit: Effective and Flexible Diffusion-Based Editing via Doob's h-TransformToan Nguyen, Kien Do, Duc Kieu, Thin NguyenCVPR 2025
- Fast constrained sampling in pre-trained diffusion modelsAlexandros Graikos, Nebojsa Jojic, Dimitris SamarasNeurIPS 2025 · 被引用 9 次
- Training-Free Adaptation of Diffusion Models via Doob's -TransformQijie Zhu, Zeqi Ye, Han Liu, Zhaoran Wang 等ICML 2026 · 被引用 3 次
- Doob's Lagrangian: A Sample-Efficient Variational Approach to Transition Path SamplingYuanqi Du, Michael Plainer, Rob Brekelmans, Chenru Duan 等NeurIPS 2024 · 被引用 41 次
