Letting Trajectories Spread: Quality-Preserving Control for Diverse Flow Matching
Jingxuan Wu, Zhenglin Wan, Xingrui Yu, Yuzhe YANG, Bo An, Ivor Tsang, Yang You
Abstract
Flow-based text-to-image models follow deterministic trajectories, making it costly to explore diverse modes under limited sampling budgets. Existing approaches to improving diversity often rely on retraining or degrade image fidelity. To address this limitation, we present a training-free, inference-time control mechanism that makes the flow itself diversity-aware. Our core insight is to encourage diversity through guidance that is geometrically decoupled from the model’s quality-seeking direction. Our method simultaneously encourages lateral spread among trajectories via a feature-space objective and reintroduces uncertainty through a time-scheduled stochastic perturbation. Crucially, this perturbation is projected to be orthogonal to the generation flow, a geometric constraint that allows it to boost variation without degrading image details or prompt fidelity. Theoretically, we show that this design monotonically increases a volume surrogate while approximately preserving the marginal distribution, providing a principled explanation for the robustness of generation quality. Empirically, across multiple text-to-image settings under fixed sampling budgets, our method consistently improves diversity metrics such as the Vendi Score and Brisque over strong baselines, while upholding image quality and alignment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 3,959 citations
Related papers
- It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion ModelsAnne Harrington, A. Sophia Koepke, Shyamgopal Karthik, Trevor Darrell et al.CVPR 2026 · 12 citations
- DiverseFlow: Sample-Efficient Diverse Mode Coverage in FlowsMashrur Mahmud Morshed, Vishnu BoddetiCVPR 2025
- GASS: Geometry-Aware Spherical Sampling for Disentangled Diversity Enhancement in Text-to-Image GenerationYe Zhu, Kaleb Newman, Johannes Lutzeyer, Adriana Romero-Soriano et al.ICML 2026 · 1 citation
- Breaking the Lock-in: Diversifying Text-to-Image Generation via Representation ModulationDahee Kwon, Haeun Lee, Jaesik ChoiICML 2026
- Initialization is Half the Battle: Generating Diverse Images from a Guidance Potential PosteriorXiang Li, Dianbo Liu, Kenji KawaguchiICML 2026
