Self-Supervised Learning with an Information Maximization Criterion
Serdar Ozsoy, Shadi Hamdan, Sercan Ö. Arik, Deniz Yuret, Alper T. Erdogan
Abstract
Self-supervised learning allows AI systems to learn effective representations from large amounts of data using tasks that do not require costly labeling. Mode collapse, i.e., the model producing identical representations for all inputs, is a central problem to many self-supervised learning approaches, making self-supervised tasks, such as matching distorted variants of the inputs, ineffective. In this article, we argue that a straightforward application of information maximization among alternative latent representations of the same input naturally solves the collapse problem and achieves competitive empirical results. We propose a self-supervised learning method, CorInfoMax, that uses a second-order statistics-based mutual information measure that reflects the level of correlation among its arguments. Maximizing this correlative information measure between alternative representations of the same input serves two purposes: (1) it avoids the collapse problem by generating feature vectors with non-degenerate covariances; (2) it establishes relevance among alternative representations by increasing the linear dependence among them. An approximation of the proposed information maximization objective simplifies to a Euclidean distance-based objective function regularized by the log-determinant of the feature covariance matrix. The regularization term acts as a natural barrier against feature space degeneracy. Consequently, beyond avoiding complete output collapse to a single point, the proposed approach also prevents dimensional collapse by encouraging the spread of information across the whole feature space. Numerical experiments demonstrate that CorInfoMax achieves better or competitive performance results relative to the state-of-the-art SSL approaches. Preprint. Under review.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 51c8c8c3-bd3a-4e73-af8c-58ba3cba2c45Cited by top-tier papers13
- Learning Efficient Coding of Natural Images with Maximum Manifold Capacity RepresentationsThomas E. Yerxa, Yilun Kuang, Eero P. Simoncelli, SueYeon ChungNeurIPS 2023 · 44 citations
- Self-Supervised Learning of Representations for Space Generates Multi-Modular Grid CellsRylan Schaeffer, Mikail Khona, Tzuhsuan Ma, Cristóbal Eyzaguirre et al.NeurIPS 2023 · 40 citations
- Correlative Information Maximization: A Biologically Plausible Approach to Supervised Deep Neural Networks without Weight SymmetryBariscan Bozkurt, Cengiz Pehlevan, Alper T. ErdoganNeurIPS 2023 · 6 citations
- Contrastive Self-Supervised Learning As Neural Manifold PackingGuanming Zhang, David J. Heeger, Stefano MartinianiNeurIPS 2025 · 5 citations
- Error Broadcast and Decorrelation as a Potential Artificial and Natural Learning MechanismMete Erdogan, Cengiz Pehlevan, Alper T. ErdoganNeurIPS 2025 · 3 citations
Builds on18
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 1,226 citations
Related papers
- Preventing Dimensional Collapse in Self-Supervised Learning via Orthogonality RegularizationJunlin He, Jinxiao Du, Wei MaNeurIPS 2024 · 19 citations
- On Feature Decorrelation in Self-Supervised LearningTianyu Hua, Wenxiao Wang, Zihui Xue, Sucheng Ren et al.ICCV 2021 · 237 citations
- An Information Theory Perspective on Variance-Invariance-Covariance RegularizationRavid Shwartz-Ziv, Randall Balestriero, Kenji Kawaguchi, Tim G. J. Rudner et al.NeurIPS 2023 · 21 citations
- Geometric View of Soft Decorrelation in Self-Supervised LearningYifei Zhang, Hao Zhu, Zixing Song, Yankai Chen et al.KDD 2024 · 15 citations
- Rethinking the Uniformity Metric in Self-Supervised LearningXianghong Fang, Jian Li, Qiang Sun, Benyou WangICLR 2024 · 3 citations
