Investigating Why Contrastive Learning Benefits Robustness against Label Noise
Yihao Xue, Kyle Whitecross, Baharan Mirzasoleiman
Abstract
Self-supervised Contrastive Learning (CL) has been recently shown to be very effective in preventing deep networks from overfitting noisy labels. Despite its empirical success, the theoretical understanding of the effect of contrastive learning on boosting robustness is very limited. In this work, we rigorously prove that the representation matrix learned by contrastive learning boosts robustness, by having: (i) one prominent singular value corresponding to each sub-class in the data, and significantly smaller remaining singular values; and (ii) a large alignment between the prominent singular vectors and the clean labels of each sub-class. The above properties enable a linear layer trained on such representations to effectively learn the clean labels without overfitting the noise. We further show that the low-rank structure of the Jacobian of deep networks pre-trained with contrastive learning allows them to achieve a superior performance initially, when fine-tuned on noisy labels. Finally, we demonstrate that the initial robustness provided by contrastive learning enables robust training methods to achieve state-of-the-art performance under extreme noise levels, e.g., an average of 27.18% and 15.58% increase in accuracy on CIFAR-10 and CIFAR-100 with 80% symmetric noisy labels, and 4.11% increase in accuracy on WebVision.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ef7b3819-1246-4176-91ef-e04344f89931Cited by top-tier papers19
- FOCAL: Contrastive Learning for Multimodal Time-Series Sensing Signals in Factorized Orthogonal Latent SpaceShengzhong Liu, Tomoyoshi Kimura, Dongxin Liu, Ruijie Wang et al.NeurIPS 2023 · 72 citations
- Understanding Contrastive Learning via Distributionally Robust OptimizationJunkang Wu, Jiawei Chen, Jiancan Wu, Wentao Shi et al.NeurIPS 2023 · 55 citations
- Understanding and Mitigating the Label Noise in Pre-training on Downstream TasksHao Chen, Jindong Wang, Ankit Shah, Ran Tao et al.ICLR 2024 · 49 citations
- When Noisy Labels Meet Long Tail Dilemmas: A Representation Calibration MethodManyi Zhang, Xuyang Zhao, Jun Yao, Chun Yuan et al.ICCV 2023 · 37 citations
- Group Robust Classification Without Any Group InformationChristos Tsirigotis, João Monteiro, Pau Rodríguez, David Vázquez et al.NeurIPS 2023 · 34 citations
Builds on6
- Early-Learning Regularization Prevents Memorization of Noisy LabelsSheng Liu, Jonathan Niles-Weed, Narges Razavian, Carlos Fernandez-GrandaNeurIPS 2020 · 798 citations
- Provable Guarantees for Self-Supervised Deep Learning with Spectral Contrastive LossJeff Z. HaoChen, Colin Wei, Adrien Gaidon, Tengyu MaNeurIPS 2021 · 425 citations
- Coresets for Robust Training of Deep Neural Networks against Noisy LabelsBaharan Mirzasoleiman, Kaidi Cao, Jure LeskovecNeurIPS 2020 · 99 citations
- The Surprising Simplicity of the Early-Time Learning Dynamics of Neural NetworksWei Hu, Lechao Xiao, Ben Adlam, Jeffrey PenningtonNeurIPS 2020 · 77 citations
- Heteroskedastic and Imbalanced Deep Learning with Adaptive RegularizationKaidi Cao, Yining Chen, Junwei Lu, Nikos Aréchiga et al.ICLR 2021 · 20 citations
Related papers
- Selective-Supervised Contrastive Learning with Noisy LabelsShikun Li, Xiaobo Xia, Shiming Ge, Tongliang LiuCVPR 2022 · 201 citations
- On Learning Contrastive Representations for Learning with Noisy LabelsLi Yi, Sheng Liu, Qi She, A. Ian McLeod et al.CVPR 2022 · 68 citations
- When does Contrastive Learning Preserve Adversarial Robustness from Pretraining to Finetuning?Lijie Fan, Sijia Liu, Pin-Yu Chen, Gaoyuan Zhang et al.NeurIPS 2021 · 147 citations
- Scarf: Self-Supervised Contrastive Learning using Random Feature CorruptionDara Bahri, Heinrich Jiang, Yi Tay, Donald MetzlerICLR 2022 · 233 citations
- Adversarial Self-Supervised Contrastive LearningMinseon Kim, Jihoon Tack, Sung Ju HwangNeurIPS 2020 · 294 citations
