Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment
Johannes Schusterbauer, Ming Gui, Frank Fundel, Björn Ommer
摘要
Diffusion models have revolutionized generative tasks through high-fidelity outputs, yet flow matching (FM) offers faster inference and empirical performance gains. However, current foundation FM models are computationally prohibitive for finetuning, while diffusion models like Stable Diffusion benefit from efficient architectures and ecosystem support. This work addresses the critical challenge of efficiently transferring knowledge from pre-trained diffusion models to flow matching. We propose Diff2Flow, a novel framework that systematically bridges diffusion and FM paradigms by rescaling timesteps, aligning interpolants, and deriving FM-compatible velocity fields from diffusion predictions. This alignment enables direct and efficient FM finetuning of diffusion priors with no extra computation overhead. Our experiments demonstrate that Diff2Flow outperforms naïve FM and diffusion finetuning particularly under parameter-efficient constraints, while achieving superior or competitive performance across diverse downstream tasks compared to state-of-the-art methods. We will release our code at https://github. com/CompVis/diff2flow.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Adapting Self-Supervised Representations as a Latent Space for Efficient GenerationMing Gui, Johannes Schusterbauer, Timy Phan, Felix Krause 等ICLR 2026 · 被引用 13 次
- RMFlow: Refined Mean Flow by a Noise-Injection Step for Multimodal GenerationYuhao Huang, Shih-Hsin Wang, Andrea L. Bertozzi, Bao WangICLR 2026 · 被引用 6 次
- FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion DeblurringXiaoyang Liu, Zhengyan Zhou, Zihang Xu, Jiezhang Cao 等ICLR 2026 · 被引用 5 次
- Gaussian Mixture Flow Matching with Domain Alignment for Multi-Domain Sequential RecommendationXiaoxin Ye, Chengkai Huang, Hongtao Huang, Lina YaoWWW 2026 · 被引用 4 次
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li 等CVPR 2026 · 被引用 4 次
它引用的顶会 Paper33
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
相关 Paper
- Value Gradient Guidance for Flow Matching AlignmentZhen Liu, Tim Z. Xiao, Carles Domingo-Enrich, Weiyang Liu 等NeurIPS 2025 · 被引用 15 次
- BiFM: Bidirectional Flow Matching for Few-Step Image Editing and GenerationYasong Dai, Zeeshan Hayder, David Ahmedt-Aristizabal, Hongdong LiCVPR 2026 · 被引用 1 次
- 2ndMatch: Finetuning Pruned Diffusion Models via Second-Order Jacobian MatchingCaleb Zheng, Eli ShlizermanCVPR 2026
- Rectified Diffusion: Straightness Is Not Your Need in Rectified FlowFu-Yun Wang, Ling Yang, Zhaoyang Huang, Mengdi Wang 等ICLR 2025
- FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image EditingJeongsol Kim, Yeobin Hong, Jonghyun Park, Jong Chul YeICLR 2026 · 被引用 35 次
