-ReQ : Assessing Representation Quality in Self-Supervised Learning by measuring eigenspectrum decay
Kumar Krishna Agrawal, Arnab Kumar Mondal, Arna Ghosh, Blake A. Richards
Abstract
Self-Supervised Learning (SSL) with large-scale unlabelled datasets enables learning useful representations for multiple downstream tasks. However, efficiently assessing the quality of such representations poses nontrivial challenges. Existing approaches train linear probes (with frozen features) to evaluate performance on a given task. This is expensive both computationally, since it requires retraining a new prediction head for each downstream task, and statistically, which requires task-specific labels for multiple tasks. This poses a natural question, how do we efficiently determine the "goodness" of representations learned with SSL across a wide range of potential downstream tasks? In particular, a task-agnostic statistical measure of representation quality that predicts generalization without explicit downstream task evaluation would be highly desirable. In this work, we analyze characteristics of learned representations f ✓ in well-trained neural networks with canonical architectures & across SSL objectives. We observe that the eigenspectrum of the empirical feature covariance Cov(f ✓ ) can be well approximated with the family of a power-law distribution. We analytically and empirically (using multiple datasets, e.g. CIFAR, STL10, MIT67, ImageNet) demonstrate that the decay coefficient ↵ serves as a measure of representation quality for tasks that are solvable with a linear readout, that is, there exist welldefined intervals for ↵ where models exhibit excellent downstream generalization. Furthermore, our experiments suggest that key design parameters in SSL algorithms, such as BarlowTwins [1], implicitly modulate the decay coefficient of the eigenspectrum (↵). As ↵ depends only on the features themselves, this measure for model selection with hyperparameter tuning for BarlowTwins enables the search with less compute.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 674e27d0-9ddc-4f71-81e4-2b9343de2305Cited by top-tier papers3
- Temperature Balancing, Layer-wise Weight Analysis, and Neural Network TrainingYefan Zhou, Tianyu Pang, Keqin Liu, Charles H. Martin et al.NeurIPS 2023 · 29 citations
- When is an Embedding Model More Promising than Another?Maxime Darrin, Philippe Formont, Ismail Ben Ayed, Jackie CK Cheung et al.NeurIPS 2024 · 10 citations
- PPGPT: Transferring Next-Token Modeling from Language to PPG SignalsZexing Zhang, Huimin Lu, Qingxin ZhaoAAAI 2026 · 1 citation
Builds on11
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- Do Vision Transformers See Like Convolutional Neural Networks?Maithra Raghu, Thomas Unterthiner, Simon Kornblith, Chiyuan Zhang et al.NeurIPS 2021 · 1,553 citations
Related papers
- IdEst: Assessing Self-Supervised Learning Representations via Intrinsic DimensionJulie Mordacq, Vicky Kalogeiton, Steve OudotICML 2026 · 1 citation
- Measuring Self-Supervised Representation Quality for Downstream Classification Using Discriminative FeaturesNeha Mukund Kalibhat, Kanika Narang, Hamed Firooz, Maziar Sanjabi et al.AAAI 2024 · 12 citations
- RankMe: Assessing the Downstream Performance of Pretrained Self-Supervised Representations by Their RankQuentin Garrido, Randall Balestriero, Laurent Najman, Yann LeCunICML 2023 · 127 citations
- Measuring the Interpretability of Unsupervised Representations via Quantized Reversed ProbingIro Laina, Yuki M. Asano, Andrea VedaldiICLR 2022 · 9 citations
- Exploring the Gap between Collapsed & Whitened Features in Self-Supervised LearningBobby He, Mete OzayICML 2022 · 31 citations
