FlowTurbo: Towards Real-time Flow-Based Image Generation with Velocity Refiner
Wenliang Zhao, Minglei Shi, Xumin Yu, Jie Zhou, Jiwen Lu
摘要
Building on the success of diffusion models in visual generation, flow-based models reemerge as another prominent family of generative models that have achieved competitive or better performance in terms of both visual quality and inference speed. By learning the velocity field through flow-matching, flow-based models tend to produce a straighter sampling trajectory, which is advantageous during the sampling process. However, unlike diffusion models for which fast samplers are well-developed, efficient sampling of flow-based generative models has been rarely explored. In this paper, we propose a framework called FlowTurbo to accelerate the sampling of flow-based models while still enhancing the sampling quality. Our primary observation is that the velocity predictor's outputs in the flow-based models will become stable during the sampling, enabling the estimation of velocity via a lightweight velocity refiner. Additionally, we introduce several techniques including a pseudo corrector and sample-aware compilation to further reduce inference time. Since FlowTurbo does not change the multi-step sampling paradigm, it can be effectively applied for various tasks such as image editing, inpainting, etc. By integrating FlowTurbo into different flow-based models, we obtain an acceleration ratio of 53.1%58.3% on class-conditional generation and 29.8%38.5% on text-to-image generation. Notably, FlowTurbo reaches an FID of 2.12 on ImageNet with 100 (ms / img) and FID of 3.93 with 38 (ms / img), achieving the real-time image generation and establishing the new state-of-the-art. Code is available at https://github.com/shiml20/FlowTurbo.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion ParadigmZiyan Guo, Zeyu Hu, De Wen Soh, Na ZhaoICCV 2025 · 被引用 10 次
- ImageDoctor: Diagnosing Text-to-Image Generation via Grounded Image ReasoningYuxiang Guo, Jiang Liu, Ze Wang, Hao Chen 等ICLR 2026 · 被引用 5 次
- Flow Matching for Denoised Social RecommendationYinxuan Huang, Ke Liang, Zhuofan Dong, Xiaodong Qu 等ICML 2025
它引用的顶会 Paper22
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- FlowCast: Trajectory Forecasting for Scalable Zero-Cost Speculative Flow MatchingDivya Jyoti Bajpai, Shubham Agarwal, Apoorv Saxena, Kuldeep Kulkarni 等ICLR 2026 · 被引用 3 次
- FastFlow: Accelerating The Generative Flow Matching Models with Bandit InferenceDivya Jyoti Bajpai, Dhruv Bhardwaj, Soumya Roy, Tejas Duseja 等ICLR 2026 · 被引用 3 次
- A-FloPS: Accelerating Diffusion Models via Adaptive Flow Path SamplerCheng Jin, Zhenyu Xiao, Yuantao GuAAAI 2026
- Deeply Supervised Flow-Based Generative ModelsInkyu Shin, Chenglin Yang, Liang-Chieh ChenICCV 2025 · 被引用 1 次
- Simple ReFlow: Improved Techniques for Fast Flow ModelsBeomsu Kim, Yu-Guan Hsieh, Michal Klein, Marco Cuturi 等ICLR 2025
