Geometric View of Soft Decorrelation in Self-Supervised Learning
Yifei Zhang, Hao Zhu, Zixing Song, Yankai Chen, Xinyu Fu, Ziqiao Meng, Piotr Koniusz, Irwin King
摘要
Contrastive learning, a form of Self-Supervised Learning (SSL), typically consists of an alignment term and a regularization term. The alignment term minimizes the distance between the embeddings of a positive pair, while the regularization term prevents trivial solutions and expresses prior beliefs about the embeddings. As a widely used regularization technique, soft decorrelation has been employed by several non-contrastive SSL methods to avoid trivial solutions. While the decorrelation term is designed to address the issue of dimensional collapse, we find that it fails to achieve this goal theoretically and experimentally. Based on such a finding, we extend the soft decorrelation regularization to minimize the distance between the covariance matrix and an identity matrix. We provide a new perspective on the geometric distance between positive definite matrices to investigate why the soft decorrelation cannot efficiently solve the dimensional collapse. Furthermore, we construct a family of loss functions utilizing the Bregman Matrix Divergence (BMD), with the soft decorrelation representing a specific instance within this family. We prove that a loss function (LogDet) in this family can solve the issue of dimensional collapse. Our novel loss functions based on BMD exhibit superior performance compared to the soft decorrelation and other baseline techniques, as demonstrated by experimental results on graph and image datasets.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper9
- Multi-Pair Temporal Sentence Grounding via Multi-Thread Knowledge Transfer NetworkXiang Fang, Wanlong Fang, Changshuo Wang, Daizong Liu 等AAAI 2025 · 被引用 10 次
- Astra: Efficient Transformer Architecture and Contrastive Dynamics Learning for Embodied Instruction FollowingYueen Ma, Dafeng Chi, Shiguang Wu, Yuecheng Liu 等EMNLP 2025 · 被引用 9 次
- Semi-supervised Node Importance Estimation with Informative Distribution Modeling for Uncertainty RegularizationYankai Chen, Taotao Wang, Yixiang Fang, Yunyu XiaoWWW 2025 · 被引用 8 次
- CrossSpectra: Exploiting Cross-Layer Smoothness for Parameter-Efficient Fine-TuningYifei Zhang, Hao Zhu, Junhao Dong, Haoran Shi 等NeurIPS 2025 · 被引用 5 次
- Understanding and Mitigating Hyperbolic Dimensional Collapse in Graph Contrastive LearningYifei Zhang, Hao Zhu, Menglin Yang, Jiahong Liu 等KDD 2025 · 被引用 4 次
相关 Paper
- Zero-CL: Instance and Feature decorrelation for negative-free symmetric contrastive learningShaofeng Zhang, Feng Zhu, Junchi Yan, Rui Zhao 等ICLR 2022 · 被引用 52 次
- On Feature Decorrelation in Self-Supervised LearningTianyu Hua, Wenxiao Wang, Zihui Xue, Sucheng Ren 等ICCV 2021 · 被引用 237 次
- Self-Supervised Learning with an Information Maximization CriterionSerdar Ozsoy, Shadi Hamdan, Sercan Ö. Arik, Deniz Yuret 等NeurIPS 2022 · 被引用 54 次
- GradGCL: Gradient Graph Contrastive LearningRan Li, Shimin Di, Lei Chen, Xiaofang ZhouICDE 2024 · 被引用 3 次
- Understanding Dimensional Collapse in Contrastive Self-supervised LearningLi Jing, Pascal Vincent, Yann LeCun, Yuandong TianICLR 2022 · 被引用 467 次
