Exploring the Gap between Collapsed & Whitened Features in Self-Supervised Learning
Bobby He, Mete Ozay
摘要
Avoiding feature collapse, when a Neural Network (NN) encoder maps all inputs to a constant vector, is a shared implicit desideratum of various methodological advances in self-supervised learning (SSL). To that end, whitened features have been proposed as an explicit objective to ensure uncollapsed features (Zbontar et al., 2021; Ermolov et al., 2021; Hua et al., 2021; Bardes et al., 2022) . We identify power law behaviour in eigenvalue decay, parameterised by exponent β≥0, as a spectrum that bridges between the collapsed & whitened feature extremes. We provide theoretical & empirical evidence highlighting the factors in SSL, like projection layers & regularisation strength, that influence eigenvalue decay rate, & demonstrate that the degree of feature whitening affects generalisation, particularly in label scarce regimes. We use our insights to motivate a novel method, Post-hoc Manipulation of the Principal Axes & Trace (PostMan-Pat), which efficiently post-processes a pretrained encoder to enforce eigenvalue decay rate with power law exponent β, & find that PostMan-Pat delivers improved label efficiency and transferability across a range of SSL methods and encoder architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- RankMe: Assessing the Downstream Performance of Pretrained Self-Supervised Representations by Their RankQuentin Garrido, Randall Balestriero, Laurent Najman, Yann LeCunICML 2023 · 被引用 127 次
- The SSL Interplay: Augmentations, Inductive Bias, and GeneralizationVivien Cabannes, Bobak Toussi Kiani, Randall Balestriero, Yann LeCun 等ICML 2023 · 被引用 43 次
- -ReQ : Assessing Representation Quality in Self-Supervised Learning by measuring eigenspectrum decayKumar Krishna Agrawal, Arnab Kumar Mondal, Arna Ghosh, Blake A. RichardsNeurIPS 2022 · 被引用 39 次
- Preventing Dimensional Collapse in Self-Supervised Learning via Orthogonality RegularizationJunlin He, Jinxiao Du, Wei MaNeurIPS 2024 · 被引用 19 次
- LDReg: Local Dimensionality Regularized Self-Supervised LearningHanxun Huang, Ricardo J. G. B. Campello, Sarah Monazam Erfani, Xingjun Ma 等ICLR 2024 · 被引用 12 次
它引用的顶会 Paper27
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
相关 Paper
- Modulate Your Spectrum in Self-Supervised LearningXi Weng, Yunhao Ni, Tengwei Song, Jie Luo 等ICLR 2024 · 被引用 10 次
- On Feature Decorrelation in Self-Supervised LearningTianyu Hua, Wenxiao Wang, Zihui Xue, Sucheng Ren 等ICCV 2021 · 被引用 237 次
- What shapes the loss landscape of self supervised learning?Liu Ziyin, Ekdeep Singh Lubana, Masahito Ueda, Hidenori TanakaICLR 2023 · 被引用 2 次
- An Investigation into Whitening Loss for Self-supervised LearningXi Weng, Lei Huang, Lei Zhao, Rao Muhammad Anwer 等NeurIPS 2022 · 被引用 26 次
- A Random Matrix Theory of Masked Self-Supervised LearningArie Zurich, Federica Gerace, Bruno Loureiro, Yue LuICML 2026
