Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps
Peter Holderrieth, Douglas Chen, Luca Eyring, Ishin Shah, Giri Anantharaman, Yutong He, Zeynep Akata, Tommi Jaakkola, Nicholas Boffi, Max Simchowitz
Abstract
Flow and diffusion models produce high-quality samples, but adapting them to user preferences or constraints post-training remains costly and brittle, a challenge commonly called reward alignment. We argue that efficient reward alignment should be a property of the generative model itself, not an afterthought, and redesign the model for adaptability. We propose Diamond Maps, a stochastic flow-map model that enables efficient and accurate alignment to arbitrary rewards at inference time. Diamond Maps amortize many simulation steps into a single-step sampler, like flow maps, while preserving the stochasticity required for optimal reward adaptation. This design makes search, Sequential Monte Carlo, and guidance scalable by enabling efficient and consistent estimation of the value function. Our experiments show that Diamond Maps can be learned efficiently via distillation from GLASS Flows, achieve stronger reward-alignment performance, and scale better than existing alignment methods. Overall, our results point toward a practical route to generative models that can be rapidly adapted to arbitrary preferences and constraints at inference time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8bae1e3c-1deb-46f0-b10f-f2bdf05fecbcCited by top-tier papers2
- Meta Flow Maps enable scalable reward alignmentPeter Potaptchik, Adhi Saravanan, Abbas Mammadov, Alvaro Prat et al.ICML 2026 · 25 citations
- How to Guide Your Flow: Few-Step Alignment via Flow Map Reward GuidanceJerry Huang, Justin Lin, Sheel Shah, Kartik Nair et al.ICML 2026
Builds on43
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari et al.ICML 2024 · 3,620 citations
- Consistency ModelsYang Song, Prafulla Dhariwal, Mark Chen, Ilya SutskeverICML 2023 · 1,720 citations
Related papers
- Diffusion Tree Sampling: Scalable inference‑time alignment of diffusion modelsVineet Jain, Kusha Sareen, Mohammad Pedramfar, Siamak RavanbakhshNeurIPS 2025 · 41 citations
- Test-time Alignment of Diffusion Models without Reward Over-optimizationSunwoo Kim, Minkyu Kim, Dongmin ParkICLR 2025
- GLASS Flows: Efficient Inference for Reward Alignment of Flow and Diffusion ModelsPeter Holderrieth, Uriel Singer, Tommi Jaakkola, Ricky T. Q. Chen et al.ICLR 2026 · 7 citations
- Categorical Flow MapsDaan Roos, Oscar Davis, Floor Eijkelboom, Michael Bronstein et al.ICML 2026 · 23 citations
- Nabla-R2D3: Effective and Efficient 3D Diffusion Alignment with 2D RewardsQingming Liu, Zhen Liu, Dinghuai Zhang, Kui JiaNeurIPS 2025 · 10 citations
