Contrastive Flow Matching
George Stoica, Vivek Ramanujan, Xiang Fan, Ali Farhadi, Ranjay Krishna, Judy Hoffman
Abstract
Unconditional flow-matching trains diffusion models to transport samples from a source distribution to a target distribution by enforcing that the flows between sample pairs are unique. However, in conditional settings (e.g., class-conditioned models), this uniqueness is no longer guaranteed--flows from different conditions may overlap, leading to more ambiguous generations. We introduce Contrastive Flow Matching, an extension to the flow matching objective that explicitly enforces uniqueness across all conditional flows, enhancing condition separation. Our approach adds a contrastive objective that maximizes dissimilarities between predicted flows from arbitrary sample pairs. We validate Contrastive Flow Matching by conducting extensive experiments across varying model architectures on both class-conditioned (ImageNet-1k) and text-to-image (CC3M) benchmarks. Notably, we find that training models with Contrastive Flow Matching (1) improves training speed by a factor of up to 9x, (2) requires up to 5x fewer de-noising steps and (3) lowers FID by up to 8.9 compared to training the same models with flow matching. We release our code at: https://github.com/gstoica27/DeltaFM.git.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9ede20f0-ea76-49a2-99d7-1e912463a1b2Cited by top-tier papers11
- Latent Diffusion Model without Variational AutoencoderMinglei Shi, Haolin Wang, Wenzhao Zheng, Ziyang Yuan et al.ICLR 2026 · 85 citations
- There is No VAE: End-to-End Pixel-Space Generative Modeling via Self-Supervised Pre-TrainingJiachen Lei, Keli Liu, Julius Berner, Y HoiM et al.ICLR 2026 · 24 citations
- Unified Multi-Modal Interactive and Reactive 3D Motion Generation via Rectified FlowPrerit Gupta, Shourya Verma, Ananth Grama, Aniket BeraICLR 2026 · 7 citations
- WeMMU: Enhanced Bridging of Vision-Language Models and Diffusion Models via Noisy Query TokensJian Yang, Dacheng Yin, Xiaoxuan He, Yong Li et al.CVPR 2026 · 1 citation
- Align Your Trajectory Tangent: Training Better Consistency Models via Manifold-Aligned TangentsBeomsu Kim, ByungHee Cha, Jong Chul YEICML 2026 · 1 citation
Builds on25
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Contrastive Flow Map MatchingJunyu Zhang, Daochang Liu, Younghyun Kim, Jong Hwan Ko et al.ICML 2026
- Discrete Contrastive Diffusion for Cross-Modal Music and Image GenerationYe Zhu, Yu Wu, Kyle Olszewski, Jian Ren et al.ICLR 2023 · 10 citations
- Consistent Diffusion Models: Mitigating Sampling Drift by Learning to be ConsistentGiannis Daras, Yuval Dagan, Alex Dimakis, Constantinos DaskalakisNeurIPS 2023 · 79 citations
- Flow Matching for Multimodal DistributionsGaoxiang Luo, Frank Cole, Sihang Zhang, Yuxiang Wan et al.CVPR 2026 · 1 citation
- Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality EvolutionQihao Liu, Xi Yin, Alan L. Yuille, Andrew Brown et al.CVPR 2025
