Inference-Time Scaling for Flow Models via Stochastic Generation and Rollover Budget Forcing
Jaihoon Kim, Taehoon Yoon, Jisung Hwang, Minhyuk Sung
摘要
We propose an inference-time scaling approach for pretrained flow models. Recently, inference-time scaling has gained significant attention in LLMs and diffusion models, improving sample quality or better aligning outputs with user preferences by leveraging additional computation. For diffusion models, particle sampling has allowed more efficient scaling due to the stochasticity at intermediate denoising steps. On the contrary, while flow models have gained popularity as an alternative to diffusion models--offering faster generation and high-quality outputs in state-of-the-art image and video generative models--efficient inference-time scaling methods used for diffusion models cannot be directly applied due to their deterministic generative process. To enable efficient inference-time scaling for flow models, we propose three key ideas: 1) SDE-based generation, enabling particle sampling in flow models, 2) Interpolant conversion, broadening the search space and enhancing sample diversity, and 3) Rollover Budget Forcing (RBF), an adaptive allocation of computational resources across timesteps to maximize budget utilization. Our experiments show that SDE-based generation, particularly variance-preserving (VP) interpolant-based generation, improves the performance of particle sampling methods for inference-time scaling in flow models. Additionally, we demonstrate that RBF with VP-SDE achieves the best performance, outperforming all previous inference-time scaling approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Flow-GRPO: Training Flow Matching Models via Online RLJie Liu, Gongye Liu, Jiajun Liang, Yangguang Li 等NeurIPS 2025 · 被引用 647 次
- Test-Time Scaling of Diffusion Models via Noise Trajectory SearchVignav Ramesh, Morteza MardaniNeurIPS 2025 · 被引用 33 次
- BézierFlow: Learning Bézier Stochastic Interpolant Schedulers for Few-Step GenerationYunhong Min, Juil Koo, Seungwoo Yoo, Minhyuk SungICLR 2026 · 被引用 4 次
- Rethinking Prompt Design for Inference-time Scaling in Text-to-Visual GenerationSubin Kim, Sangwoo Mo, Mamshad Nayeem Rizve, Yiran Xu 等CVPR 2026 · 被引用 2 次
- Diffusion Alignment as Variational Expectation-MaximizationJaewoo Lee, Minsu Kim, Sanghyeok Choi, Inhyuck Song 等ICLR 2026 · 被引用 2 次
它引用的顶会 Paper43
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- IV-mixed Sampler: Leveraging Image Diffusion Models for Enhanced Video SynthesisShitong Shao, Zikai Zhou, Bai Lichen, Haoyi Xiong 等ICLR 2025
- Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative CompressionJung Yi, Wooseok Jang, Paul Cho, Jisu Nam 等ICML 2026
- XYZFlow: Scaling Multidimensional Shortcut Flows for Efficient Generative ModelingJinxiu Liu, Xuanming Liu, Kangfu Mei, Yandong Wen 等ICML 2026
- Entropic Time Schedulers for Generative Diffusion ModelsDejan Stancevic, Florian Handke, Luca AmbrogioniNeurIPS 2025 · 被引用 19 次
- Efficient Parallel Samplers for Recurrent-Depth ModelsJonas Geiping, Xinyu Yang, Guinan SuICML 2026 · 被引用 5 次
