Value Gradient Guidance for Flow Matching Alignment
Zhen Liu, Tim Z. Xiao, Carles Domingo-Enrich, Weiyang Liu, Dinghuai Zhang
Abstract
While methods exist for aligning flow matching models--a popular and effective class of generative models--with human preferences, existing approaches fail to achieve both adaptation efficiency and probabilistically sound prior preservation. In this work, we leverage the theory of optimal control and propose VGG-Flow, a gradient-matching-based method for finetuning pretrained flow matching models. The key idea behind this algorithm is that the optimal difference between the finetuned velocity field and the pretrained one should be matched with the gradient field of a value function. This method not only incorporates first-order information from the reward model but also benefits from heuristic initialization of the value function to enable fast adaptation. Empirically, we show on a popular text-to-image flow matching model, Stable Diffusion 3, that our method can finetune flow matching models under limited computational budgets while achieving effective and prior-preserving alignment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext edbef009-82e4-46cc-8d25-213178b2fdc3Cited by top-tier papers4
- Discrete Adjoint Schrödinger Bridge SamplerWei Guo, Yuchen Zhu, Xiaochen Du, Juno Nam et al.ICML 2026 · 3 citations
- Q-Flow: Stable and Expressive Reinforcement Learning with Flow-based PolicyJaeHyeok Doo, Byeongguk Jeon, Seonghyeon Ye, Kimin Lee et al.ICML 2026 · 1 citation
- Solving Inverse Problems with Flow-based Models via Model Predictive ControlGeorge Webber, Alexander Denker, Riccardo Barbano, Andrew ReaderICML 2026 · 1 citation
- Conflict-Aware Additive Guidance for Flow Models under Compositional RewardsXuehui Yu, Fucheng Cai, Meiyi Wang, Xiaopeng Fan et al.ICML 2026
Builds on42
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari et al.ICML 2024 · 3,620 citations
- ImageReward: Learning and Evaluating Human Preferences for Text-to-Image GenerationJiazheng Xu, Xiao Liu, Yuchen Wu, Yuxuan Tong et al.NeurIPS 2023 · 1,310 citations
Related papers
- LeapAlign: Post-training Flow Matching Models at Any Generation Step by Building Two-Step TrajectoriesZhanhao Liang, Tao Yang, Jie Wu, Chengjian Feng et al.CVPR 2026 · 6 citations
- Diff2Flow: Training Flow Matching Models via Diffusion Model AlignmentJohannes Schusterbauer, Ming Gui, Frank Fundel, Björn OmmerCVPR 2025
- Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNetsZhen Liu, Tim Z. Xiao, Weiyang Liu, Yoshua Bengio et al.ICLR 2025
- PC-Flow: Preference Alignment in Flow Matching via ClassifierShaomeng Wang, He Wang, Longquan Dai, Jinhui TangAAAI 2026
- Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal ControlCarles Domingo-Enrich, Michal Drozdzal, Brian Karrer, Ricky T. Q. ChenICLR 2025 · 2 citations
