Greed is Good: A Unifying Perspective on Guided Generation
Zander Blasingame, Chen Liu
摘要
Training-free guided generation is a widely used and powerful technique that allows the end user to exert further control over the generative process of flow/diffusion models. Generally speaking, two families of techniques have emerged for solving this problem for gradient-based guidance: namely, posterior guidance (i.e., guidance by projecting the current sample to the target distribution via the target prediction model) and end-to-end guidance (i.e., guidance by performing backpropagation throughout the entire ODE solve). In this work, we show that these two seemingly separate families can actually be unified by looking at the posterior guidance as a greedy strategy of end-to-end guidance. We explore the theoretical connections between these two families and provide an in-depth theoretical understanding of these two techniques relative to the continuous ideal gradients. Motivated by this analysis, we then show a method for interpolating between these two families enabling a trade-off between compute and accuracy of the guidance gradients. We then validate this work on several inverse image problems and property-guided molecular generation. scheme (cf . Equation ( 2)) is part of the computation graph of the model reverse-mode automatic differentiation (Linnainmaa 1976) is applied, i.e., vanilla backpropagation. The memory cost of such techniques, however, is O(n), prompting researchers to explore the second method known as optimize-then-discretize (OTD) which instead solves another ODE in reverse-time which models the continuous-time dynamics of reverse-mode differentiation, this is called the continuous adjoint method (R. T. Chen et al. 2018; cf . Kidger 2022, Section 5.1.2).
Given a flow model u θ ∈ C 1,1 ([0, 1] × R d ; R d ) that is Lipschitz continuous in its second argument and the solution x : [0, 1] → R d , x t → x(t), let a x := ∂L/∂x t denote the adjoint state. Then a x (t) can be found by solving the continuous adjoint equation:
N.B., this technique was first proposed by Pontryagin et al. (1963) and popularized for neural differential equations by R. T. Chen et al. (2018). This approach has a constant memory cost O(1); however, this comes with the cost of several drawbacks related to the numerical scheme. While these issues are not particularly relevant to our theoretical analyses, we note them in Appendix E for the ML practitioner.
Now returning back to our problem statement from Equation ( 3), the end-to-end guidance techniques amount to optimizing the initial condition x 0 in light of the entire solution trajectory admitted by the numerical scheme. A natural question we consider for problems of this form is that rather than finding the full sequence x n , can we make use of local information instead? I.e.,
Rather than solving the full ODE from x t , what if we greedily took a locally optimal step at each x t instead?
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- SAFETY-GUIDED FLOW (SGF): A UNIFIED FRAMEWORK FOR NEGATIVE GUIDANCE IN SAFE GENERATIONMingyu Kim, Young-Heon Kim, Mijung ParkICLR 2026 · 被引用 5 次
- Rex: A Family of Reversible Exponential (Stochastic) Runge-Kutta SolversZander Blasingame, Chen LiuICML 2026 · 被引用 2 次
它引用的顶会 Paper46
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- AdjointDEIS: Efficient Gradients for Diffusion ModelsZander W. Blasingame, Chen LiuNeurIPS 2024 · 被引用 8 次
- Solving Inverse Problems with Flow-based Models via Model Predictive ControlGeorge Webber, Alexander Denker, Riccardo Barbano, Andrew ReaderICML 2026 · 被引用 1 次
- Training-Free Reward-Guided Image Editing via Trajectory Optimal ControlJinho Chang, Jaemin Kim, Jong Chul YeICLR 2026 · 被引用 2 次
- FlowGrad: Controlling the Output of Generative ODEs with GradientsXingchao Liu, Lemeng Wu, Shujian Zhang, Chengyue Gong 等CVPR 2023
- One step further with Monte-Carlo sampler to guide diffusion betterMinsi Ren, Wenhao Deng, Ruiqi Feng, Tailin WuICLR 2026
