Understanding Deep Contrastive Learning via Coordinate-wise Optimization
Yuandong Tian
摘要
We show that Contrastive Learning (CL) under a broad family of loss functions (including InfoNCE) has a unified formulation of coordinate-wise optimization on the network parameter and pairwise importance , where the max player learns representation for contrastiveness, and the min player puts more weights on pairs of distinct samples that share similar representations. The resulting formulation, called -CL, unifies not only various existing contrastive losses, which differ by how sample-pair importance is constructed, but also is able to extrapolate to give novel contrastive losses beyond popular ones, opening a new avenue of contrastive loss design. These novel losses yield comparable (or better) performance on CIFAR10, STL-10 and CIFAR-100 than classic InfoNCE. Furthermore, we also analyze the max player in detail: we prove that with fixed , max player is equivalent to Principal Component Analysis (PCA) for deep linear network, and almost all local minima are global and rank-1, recovering optimal PCA solutions. Finally, we extend our analysis on max player to 2-layer ReLU networks, showing that its fixed points can have higher ranks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Understanding Contrastive Learning via Distributionally Robust OptimizationJunkang Wu, Jiawei Chen, Jiancan Wu, Wentao Shi 等NeurIPS 2023 · 被引用 55 次
- Contrastive Learning is Spectral Clustering on Similarity GraphZhiquan Tan, Yifan Zhang, Jingqin Yang, Yang YuanICLR 2024 · 被引用 34 次
- On the Comparison between Multi-modal and Single-modal Contrastive LearningWei Huang, Andi Han, Yongqiang Chen, Yuan Cao 等NeurIPS 2024 · 被引用 26 次
- Contrastive Predict-and-Search for Mixed Integer Linear ProgramsTaoan Huang, Aaron M. Ferber, Arman Zharmagambetov, Yuandong Tian 等ICML 2024 · 被引用 23 次
- Navigating the Effect of Parametrization for Dimensionality ReductionHaiyang Huang, Yingfan Wang, Cynthia RudinNeurIPS 2024 · 被引用 8 次
它引用的顶会 Paper14
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- On Layer Normalization in the Transformer ArchitectureRuibin Xiong, Yunchang Yang, Di He, Kai Zheng 等ICML 2020 · 被引用 1,388 次
相关 Paper
- Understanding and Generalizing Contrastive Learning from the Inverse Optimal Transport PerspectiveLiangliang Shi, Gu Zhang, Haoyu Zhen, Jintao Fan 等ICML 2023 · 被引用 25 次
- Understanding Contrastive Learning via Gaussian Mixture ModelsParikshit Bansal, Ali Kavis, Sujay SanghaviNeurIPS 2025 · 被引用 6 次
- Towards a Unified Framework of Contrastive Learning for Disentangled RepresentationsStefan Matthes, Zhiwei Han, Hao ShenNeurIPS 2023 · 被引用 17 次
- Empowering Collaborative Filtering with Principled Adversarial Contrastive LossAn Zhang, Leheng Sheng, Zhibo Cai, Xiang Wang 等NeurIPS 2023 · 被引用 56 次
- Unbiased Supervised Contrastive LearningCarlo Alberto Barbano, Benoit Dufumier, Enzo Tartaglione, Marco Grangetto 等ICLR 2023 · 被引用 4 次
