Exploring the Equivalence of Siamese Self-Supervised Learning via A Unified Gradient Framework
Chenxin Tao, Honghui Wang, Xizhou Zhu, Jiahua Dong, Shiji Song, Gao Huang, Jifeng Dai
摘要
Self-supervised learning has shown its great potential to extract powerful visual representations without human annotations. Various works are proposed to deal with selfsupervised learning from different perspectives: (1) contrastive learning methods (e.g., MoCo, SimCLR) utilize both positive and negative samples to guide the training direction; (2) asymmetric network methods (e.g., BYOL, Sim-Siam) get rid of negative samples via the introduction of a predictor network and the stop-gradient operation; (3) feature decorrelation methods (e.g., Barlow Twins, VICReg) instead aim to reduce the redundancy between feature dimensions. These methods appear to be quite different in the designed loss functions from various motivations. The final accuracy numbers also vary, where different networks and tricks are utilized in different works. In this work, we demonstrate that these methods can be unified into the same form. Instead of comparing their loss functions, we derive a unified formula through gradient analysis. Furthermore, we conduct fair and detailed experiments to compare their performances. It turns out that there is little gap between these methods, and the use of momentum encoder is the key factor to boost performance. From this unified framework, we propose UniGrad, a simple but effective gradient form for self-supervised learning. It does not require a memory bank or a predictor network, but can still achieve state-of-the-art performance and easily adopt other training strategies. Extensive experiments on linear evaluation and many downstream tasks also show its effectiveness. Code is released at https: //github.com/fundamentalvision/UniGrad . * Equal contribution. † This work is done when Chenxin Tao, Honghui Wang, and Jiahua Dong are interns at SenseTime Research.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper26
- Self-Supervised Learning via Maximum Entropy CodingXin Liu, Zhongdao Wang, Yali Li, Shengjin WangNeurIPS 2022 · 被引用 65 次
- ImGCL: Revisiting Graph Contrastive Learning on Imbalanced Node ClassificationLiang Zeng, Lanqing Li, Ziqi Gao, Peilin Zhao 等AAAI 2023 · 被引用 55 次
- Downstream-agnostic Adversarial ExamplesZiqi Zhou, Shengshan Hu, Ruizhi Zhao, Qian Wang 等ICCV 2023 · 被引用 45 次
- Representation Uncertainty in Self-Supervised Learning as Variational InferenceHiroki Nakamura, Masashi Okada, Tadahiro TaniguchiICCV 2023 · 被引用 27 次
- An Investigation into Whitening Loss for Self-supervised LearningXi Weng, Lei Huang, Lei Zhao, Rao Muhammad Anwer 等NeurIPS 2022 · 被引用 26 次
它引用的顶会 Paper19
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
相关 Paper
- Bridging the Gap from Asymmetry Tricks to Decorrelation Principles in Non-contrastive Self-supervised LearningKang-Jun Liu, Masanori Suganuma, Takayuki OkataniNeurIPS 2022 · 被引用 16 次
- Implicit Contrastive Representation Learning with Guided Stop-gradientByeongchan Lee, Sehyun LeeNeurIPS 2023 · 被引用 3 次
- On the duality between contrastive and non-contrastive self-supervised learningQuentin Garrido, Yubei Chen, Adrien Bardes, Laurent Najman 等ICLR 2023 · 被引用 24 次
- Crafting Better Contrastive Views for Siamese Representation LearningXiangyu Peng, Kai Wang, Zheng Zhu, Mang Wang 等CVPR 2022 · 被引用 107 次
- How Does SimSiam Avoid Collapse Without Negative Samples? A Unified Understanding with Self-supervised Contrastive LearningChaoning Zhang, Kang Zhang, Chenshuang Zhang, Trung X. Pham 等ICLR 2022 · 被引用 88 次
