Intriguing Properties of Contrastive Losses
Ting Chen, Calvin Luo, Lala Li
摘要
Contrastive loss and its variants have become very popular recently for learning visual representations without supervision. In this work, we first generalize the standard contrastive loss based on cross entropy to a broader family of losses that share an abstract form of , where hidden representations are encouraged to (1) be aligned under some transformations/augmentations, and (2) match a prior distribution of high entropy. We show that various instantiations of the generalized loss perform similarly under the presence of a multi-layer non-linear projection head, and the temperature scaling () widely used in the standard contrastive loss is (within a range) inversely related to the weighting () between the two loss terms. We then study an intriguing phenomenon of feature suppression among competing features shared acros augmented views, such as "color distribution" vs "object class". We construct datasets with explicit and controllable competing features, and show that, for contrastive learning, a few bits of easy-to-learn shared features could suppress, and even fully prevent, the learning of other sets of competing features. Interestingly, this characteristic is much less detrimental in autoencoders based on a reconstruction loss. Existing contrastive learning methods critically rely on data augmentation to favor certain sets of features than others, while one may wish that a network would learn all competing features as much as its capacity allows.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper61
- A Geometric Analysis of Neural Collapse with Unconstrained FeaturesZhihui Zhu, Tianyu Ding, Jinxin Zhou, Xiao Li 等NeurIPS 2021 · 被引用 303 次
- Can contrastive learning avoid shortcut solutions?Joshua Robinson, Li Sun, Ke Yu, Kayhan Batmanghelich 等NeurIPS 2021 · 被引用 185 次
- Rethinking Semi-Supervised Medical Image Segmentation: A Variance-Reduction PerspectiveChenyu You, Weicheng Dai, Yifei Min, Fenglin Liu 等NeurIPS 2023 · 被引用 147 次
- Understanding Contrastive Learning Requires Incorporating Inductive BiasesNikunj Saunshi, Jordan T. Ash, Surbhi Goel, Dipendra Misra 等ICML 2022 · 被引用 130 次
- Self-Supervised Learning with Kernel Dependence MaximizationYazhe Li, Roman Pogodin, Danica J. Sutherland, Arthur GrettonNeurIPS 2021 · 被引用 107 次
它引用的顶会 Paper16
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
相关 Paper
- Your contrastive learning problem is secretly a distribution alignment problemZihao Chen, Chi-Heng Lin, Ran Liu, Jingyun Xiao 等NeurIPS 2024 · 被引用 14 次
- Which Features are Learnt by Contrastive Learning? On the Role of Simplicity Bias in Class Collapse and Feature SuppressionYihao Xue, Siddharth Joshi, Eric Gan, Pin-Yu Chen 等ICML 2023 · 被引用 36 次
- Self-Weighted Contrastive Learning among Multiple Views for Mitigating Representation DegenerationJie Xu, Shuo Chen, Yazhou Ren, Xiaoshuang Shi 等NeurIPS 2023 · 被引用 71 次
- Understanding the Behaviour of Contrastive LossFeng Wang, Huaping LiuCVPR 2021
- Improving Transferability of Representations via Augmentation-Aware Self-SupervisionHankook Lee, Kibok Lee, Kimin Lee, Honglak Lee 等NeurIPS 2021 · 被引用 66 次
