Understanding Deep Contrastive Learning via Coordinate-wise Optimization
Yuandong Tian
Abstract
We show that Contrastive Learning (CL) under a broad family of loss functions (including InfoNCE) has a unified formulation of coordinate-wise optimization on the network parameter and pairwise importance , where the max player learns representation for contrastiveness, and the min player puts more weights on pairs of distinct samples that share similar representations. The resulting formulation, called -CL, unifies not only various existing contrastive losses, which differ by how sample-pair importance is constructed, but also is able to extrapolate to give novel contrastive losses beyond popular ones, opening a new avenue of contrastive loss design. These novel losses yield comparable (or better) performance on CIFAR10, STL-10 and CIFAR-100 than classic InfoNCE. Furthermore, we also analyze the max player in detail: we prove that with fixed , max player is equivalent to Principal Component Analysis (PCA) for deep linear network, and almost all local minima are global and rank-1, recovering optimal PCA solutions. Finally, we extend our analysis on max player to 2-layer ReLU networks, showing that its fixed points can have higher ranks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4eadef3c-110a-4e74-9d9e-49a20f5ddcadCited by top-tier papers21
- Understanding Contrastive Learning via Distributionally Robust OptimizationJunkang Wu, Jiawei Chen, Jiancan Wu, Wentao Shi et al.NeurIPS 2023 · 55 citations
- Contrastive Learning is Spectral Clustering on Similarity GraphZhiquan Tan, Yifan Zhang, Jingqin Yang, Yang YuanICLR 2024 · 34 citations
- On the Comparison between Multi-modal and Single-modal Contrastive LearningWei Huang, Andi Han, Yongqiang Chen, Yuan Cao et al.NeurIPS 2024 · 26 citations
- Contrastive Predict-and-Search for Mixed Integer Linear ProgramsTaoan Huang, Aaron M. Ferber, Arman Zharmagambetov, Yuandong Tian et al.ICML 2024 · 23 citations
- Navigating the Effect of Parametrization for Dimensionality ReductionHaiyang Huang, Yingfan Wang, Cynthia RudinNeurIPS 2024 · 8 citations
Builds on14
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- On Layer Normalization in the Transformer ArchitectureRuibin Xiong, Yunchang Yang, Di He, Kai Zheng et al.ICML 2020 · 1,388 citations
Related papers
- Understanding and Generalizing Contrastive Learning from the Inverse Optimal Transport PerspectiveLiangliang Shi, Gu Zhang, Haoyu Zhen, Jintao Fan et al.ICML 2023 · 25 citations
- Understanding Contrastive Learning via Gaussian Mixture ModelsParikshit Bansal, Ali Kavis, Sujay SanghaviNeurIPS 2025 · 6 citations
- Towards a Unified Framework of Contrastive Learning for Disentangled RepresentationsStefan Matthes, Zhiwei Han, Hao ShenNeurIPS 2023 · 17 citations
- Empowering Collaborative Filtering with Principled Adversarial Contrastive LossAn Zhang, Leheng Sheng, Zhibo Cai, Xiang Wang et al.NeurIPS 2023 · 56 citations
- Unbiased Supervised Contrastive LearningCarlo Alberto Barbano, Benoit Dufumier, Enzo Tartaglione, Marco Grangetto et al.ICLR 2023 · 4 citations
