Self-Supervised Learning with an Information Maximization Criterion
Serdar Ozsoy, Shadi Hamdan, Sercan Ö. Arik, Deniz Yuret, Alper T. Erdogan
摘要
Self-supervised learning allows AI systems to learn effective representations from large amounts of data using tasks that do not require costly labeling. Mode collapse, i.e., the model producing identical representations for all inputs, is a central problem to many self-supervised learning approaches, making self-supervised tasks, such as matching distorted variants of the inputs, ineffective. In this article, we argue that a straightforward application of information maximization among alternative latent representations of the same input naturally solves the collapse problem and achieves competitive empirical results. We propose a self-supervised learning method, CorInfoMax, that uses a second-order statistics-based mutual information measure that reflects the level of correlation among its arguments. Maximizing this correlative information measure between alternative representations of the same input serves two purposes: (1) it avoids the collapse problem by generating feature vectors with non-degenerate covariances; (2) it establishes relevance among alternative representations by increasing the linear dependence among them. An approximation of the proposed information maximization objective simplifies to a Euclidean distance-based objective function regularized by the log-determinant of the feature covariance matrix. The regularization term acts as a natural barrier against feature space degeneracy. Consequently, beyond avoiding complete output collapse to a single point, the proposed approach also prevents dimensional collapse by encouraging the spread of information across the whole feature space. Numerical experiments demonstrate that CorInfoMax achieves better or competitive performance results relative to the state-of-the-art SSL approaches. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Learning Efficient Coding of Natural Images with Maximum Manifold Capacity RepresentationsThomas E. Yerxa, Yilun Kuang, Eero P. Simoncelli, SueYeon ChungNeurIPS 2023 · 被引用 44 次
- Self-Supervised Learning of Representations for Space Generates Multi-Modular Grid CellsRylan Schaeffer, Mikail Khona, Tzuhsuan Ma, Cristóbal Eyzaguirre 等NeurIPS 2023 · 被引用 40 次
- Correlative Information Maximization: A Biologically Plausible Approach to Supervised Deep Neural Networks without Weight SymmetryBariscan Bozkurt, Cengiz Pehlevan, Alper T. ErdoganNeurIPS 2023 · 被引用 6 次
- Contrastive Self-Supervised Learning As Neural Manifold PackingGuanming Zhang, David J. Heeger, Stefano MartinianiNeurIPS 2025 · 被引用 5 次
- Error Broadcast and Decorrelation as a Potential Artificial and Natural Learning MechanismMete Erdogan, Cengiz Pehlevan, Alper T. ErdoganNeurIPS 2025 · 被引用 3 次
它引用的顶会 Paper18
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 被引用 1,226 次
相关 Paper
- Preventing Dimensional Collapse in Self-Supervised Learning via Orthogonality RegularizationJunlin He, Jinxiao Du, Wei MaNeurIPS 2024 · 被引用 19 次
- On Feature Decorrelation in Self-Supervised LearningTianyu Hua, Wenxiao Wang, Zihui Xue, Sucheng Ren 等ICCV 2021 · 被引用 237 次
- An Information Theory Perspective on Variance-Invariance-Covariance RegularizationRavid Shwartz-Ziv, Randall Balestriero, Kenji Kawaguchi, Tim G. J. Rudner 等NeurIPS 2023 · 被引用 21 次
- Geometric View of Soft Decorrelation in Self-Supervised LearningYifei Zhang, Hao Zhu, Zixing Song, Yankai Chen 等KDD 2024 · 被引用 15 次
- Rethinking the Uniformity Metric in Self-Supervised LearningXianghong Fang, Jian Li, Qiang Sun, Benyou WangICLR 2024 · 被引用 3 次
