DEFT: Efficient Fine-tuning of Diffusion Models by Learning the Generalised -transform
Alexander Denker, Francisco Vargas, Shreyas Padhy, Kieran Didi, Simon V. Mathis, Riccardo Barbano, Vincent Dutordoir, Emile Mathieu, Urszula Julia Komorowska, Pietro Lió
Abstract
Generative modelling paradigms based on denoising diffusion processes have emerged as a leading candidate for conditional sampling in inverse problems. In many real-world applications, we often have access to large, expensively trained unconditional diffusion models, which we aim to exploit for improving conditional sampling. Most recent approaches are motivated heuristically and lack a unifying framework, obscuring connections between them. Further, they often suffer from issues such as being very sensitive to hyperparameters, being expensive to train or needing access to weights hidden behind a closed API. In this work, we unify conditional training and sampling using the mathematically well-understood Doob's h-transform. This new perspective allows us to unify many existing methods under a common umbrella. Under this framework, we propose DEFT (Doob's h-transform Efficient FineTuning), a new approach for conditional generation that simply fine-tunes a very small network to quickly learn the conditional -transform, while keeping the larger unconditional network unchanged. DEFT is much faster than existing baselines while achieving state-of-the-art performance across a variety of linear and non-linear benchmarks. On image reconstruction tasks, we achieve speedups of up to 1.6, while having the best perceptual quality on natural images and reconstruction performance on medical images. Further, we also provide initial experiments on protein motif scaffolding and outperform reconstruction guidance methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5c57ed4e-0394-4b38-be58-9cc347aa9669Cited by top-tier papers21
- Physics-Driven Spatiotemporal Modeling for AI-Generated Video DetectionShuhai Zhang, Zihao Lian, Jiahao Yang, Daiyuan Li et al.NeurIPS 2025 · 29 citations
- Meta Flow Maps enable scalable reward alignmentPeter Potaptchik, Adhi Saravanan, Abbas Mammadov, Alvaro Prat et al.ICML 2026 · 25 citations
- RNE: plug-and-play diffusion inference-time control and energy-based trainingJiajun He, José Miguel Hernández-Lobato, Yuanqi Du, Francisco VargasICLR 2026 · 17 citations
- Value Gradient Guidance for Flow Matching AlignmentZhen Liu, Tim Z. Xiao, Carles Domingo-Enrich, Weiyang Liu et al.NeurIPS 2025 · 15 citations
- CREPE: Controlling diffusion with REPlica ExchangeJiajun He, Paul Jeha, Peter Potaptchik, Leo Zhang et al.ICLR 2026 · 11 citations
Builds on39
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
Related papers
- Supervised Guidance Training for Infinite-Dimensional Diffusion ModelsElizabeth Baker, Alexander Denker, Jes FrellsenICML 2026 · 2 citations
- h-Edit: Effective and Flexible Diffusion-Based Editing via Doob's h-TransformToan Nguyen, Kien Do, Duc Kieu, Thin NguyenCVPR 2025
- Fast constrained sampling in pre-trained diffusion modelsAlexandros Graikos, Nebojsa Jojic, Dimitris SamarasNeurIPS 2025 · 9 citations
- Training-Free Adaptation of Diffusion Models via Doob's -TransformQijie Zhu, Zeqi Ye, Han Liu, Zhaoran Wang et al.ICML 2026 · 3 citations
- Doob's Lagrangian: A Sample-Efficient Variational Approach to Transition Path SamplingYuanqi Du, Michael Plainer, Rob Brekelmans, Chenru Duan et al.NeurIPS 2024 · 41 citations
