Contrastive Flow Matching
George Stoica, Vivek Ramanujan, Xiang Fan, Ali Farhadi, Ranjay Krishna, Judy Hoffman
摘要
Unconditional flow-matching trains diffusion models to transport samples from a source distribution to a target distribution by enforcing that the flows between sample pairs are unique. However, in conditional settings (e.g., class-conditioned models), this uniqueness is no longer guaranteed--flows from different conditions may overlap, leading to more ambiguous generations. We introduce Contrastive Flow Matching, an extension to the flow matching objective that explicitly enforces uniqueness across all conditional flows, enhancing condition separation. Our approach adds a contrastive objective that maximizes dissimilarities between predicted flows from arbitrary sample pairs. We validate Contrastive Flow Matching by conducting extensive experiments across varying model architectures on both class-conditioned (ImageNet-1k) and text-to-image (CC3M) benchmarks. Notably, we find that training models with Contrastive Flow Matching (1) improves training speed by a factor of up to 9x, (2) requires up to 5x fewer de-noising steps and (3) lowers FID by up to 8.9 compared to training the same models with flow matching. We release our code at: https://github.com/gstoica27/DeltaFM.git.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Latent Diffusion Model without Variational AutoencoderMinglei Shi, Haolin Wang, Wenzhao Zheng, Ziyang Yuan 等ICLR 2026 · 被引用 85 次
- There is No VAE: End-to-End Pixel-Space Generative Modeling via Self-Supervised Pre-TrainingJiachen Lei, Keli Liu, Julius Berner, Y HoiM 等ICLR 2026 · 被引用 24 次
- Unified Multi-Modal Interactive and Reactive 3D Motion Generation via Rectified FlowPrerit Gupta, Shourya Verma, Ananth Grama, Aniket BeraICLR 2026 · 被引用 7 次
- WeMMU: Enhanced Bridging of Vision-Language Models and Diffusion Models via Noisy Query TokensJian Yang, Dacheng Yin, Xiaoxuan He, Yong Li 等CVPR 2026 · 被引用 1 次
- Align Your Trajectory Tangent: Training Better Consistency Models via Manifold-Aligned TangentsBeomsu Kim, ByungHee Cha, Jong Chul YEICML 2026 · 被引用 1 次
它引用的顶会 Paper25
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Contrastive Flow Map MatchingJunyu Zhang, Daochang Liu, Younghyun Kim, Jong Hwan Ko 等ICML 2026
- Discrete Contrastive Diffusion for Cross-Modal Music and Image GenerationYe Zhu, Yu Wu, Kyle Olszewski, Jian Ren 等ICLR 2023 · 被引用 10 次
- Consistent Diffusion Models: Mitigating Sampling Drift by Learning to be ConsistentGiannis Daras, Yuval Dagan, Alex Dimakis, Constantinos DaskalakisNeurIPS 2023 · 被引用 79 次
- Flow Matching for Multimodal DistributionsGaoxiang Luo, Frank Cole, Sihang Zhang, Yuxiang Wan 等CVPR 2026 · 被引用 1 次
- Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality EvolutionQihao Liu, Xi Yin, Alan L. Yuille, Andrew Brown 等CVPR 2025
