GULP: a prediction-based metric between representations
Enric Boix-Adserà, Hannah Lawrence, George Stepaniants, Philippe Rigollet
Abstract
Comparing the representations learned by different neural networks has recently emerged as a key tool to understand various architectures and ultimately optimize them. In this work, we introduce GULP, a family of distance measures between representations that is explicitly motivated by downstream predictive tasks. By construction, GULP provides uniform control over the difference in prediction performance between two representations, with respect to regularized linear prediction tasks. Moreover, it satisfies several desirable structural properties, such as the triangle inequality and invariance under orthogonal transformations, and thus lends itself to data embedding and visualization. We extensively evaluate GULP relative to other methods, and demonstrate that it correctly differentiates between architecture families, converges over the course of training, and captures generalization performance on downstream linear tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ec450794-1ee3-41b3-afbe-289fa72039bdCited by top-tier papers4
- Model Spider: Learning to Rank Pre-Trained Models EfficientlyYi-Kai Zhang, Ting-Ji Huang, Yao-Xiang Ding, De-Chuan Zhan et al.NeurIPS 2023 · 57 citations
- On Affine Homotopy between Language EncodersRobin Chan, Reda Boumasmoud, Anej Svete, Yuxin Ren et al.NeurIPS 2024 · 7 citations
- Towards Universality: Studying Mechanistic Similarity Across Language Model ArchitecturesJunxuan Wang, Xuyang Ge, Wentao Shu, Qiong Tang et al.ICLR 2025
- ReSi: A Comprehensive Benchmark for Representational Similarity MeasuresMax Klabunde, Tassilo Wald, Tobias Schumacher, Klaus H. Maier-Hein et al.ICLR 2025
Builds on4
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Grounding Representation Similarity Through Statistical TestingFrances Ding, Jean-Stanislas Denain, Jacob SteinhardtNeurIPS 2021 · 88 citations
- Knowledge Consistency between Neural Networks and BeyondRuofan Liang, Tianlin Li, Longfei Li, Jing Wang et al.ICLR 2020 · 30 citations
- Deconfounded Representation Similarity for Comparison of Neural NetworksTianyu Cui, Yogesh Kumar, Pekka Marttinen, Samuel KaskiNeurIPS 2022 · 27 citations
Related papers
- Learning Optimal Representations with the Decodable Information BottleneckYann Dubois, Douwe Kiela, David J. Schwab, Ramakrishna VedantamNeurIPS 2020 · 58 citations
- Generalized Shape Metrics on Neural RepresentationsAlex H. Williams, Erin Kunz, Simon Kornblith, Scott W. LindermanNeurIPS 2021 · 182 citations
- Gromov-Wasserstein AutoencodersNao Nakagawa, Ren Togo, Takahiro Ogawa, Miki HaseyamaICLR 2023 · 2 citations
- A Spectral Theory of Neural Prediction and AlignmentAbdulkadir Canatar, Jenelle Feather, Albert J. Wakhloo, SueYeon ChungNeurIPS 2023 · 29 citations
- When Does Closeness in Distribution Imply Representational Similarity? An Identifiability PerspectiveBeatrix M. G. Nielsen, Emanuele Marconato, Andrea Dittadi, Luigi GreseleNeurIPS 2025 · 7 citations
