When Representations Align: Universality in Representation Learning Dynamics
Loek van Rossem, Andrew M. Saxe
Abstract
Deep neural networks come in many sizes and architectures. The choice of architecture, in conjunction with the dataset and learning algorithm, is commonly understood to affect the learned neural representations. Yet, recent results have shown that different architectures learn representations with striking qualitative similarities. Here we derive an effective theory of representation learning under the assumption that the encoding map from input to hidden representation and the decoding map from representation to output are arbitrary smooth functions. This theory schematizes representation learning dynamics in the regime of complex, large architectures, where hidden representations are not strongly constrained by the parametrization. We show through experiments that the effective theory describes aspects of representation learning dynamics across a range of deep networks with different activation functions and architectures, and exhibits phenomena similar to the "rich" and "lazy" regime. While many network behaviors depend quantitatively on architecture, our findings point to certain behaviors that are widely conserved once models are sufficiently flexible.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Formation of Representations in Neural NetworksLiu Ziyin, Isaac L. Chuang, Tomer Galanti, Tomaso A. PoggioICLR 2025 · 1 citation
- Algorithm Development in Neural Networks: Insights from the Streaming Parity TaskLoek van Rossem, Andrew M. SaxeICML 2025
- Disentangling the Factors of Convergence between Brains and DINOv3Joséphine Raugel, Marc Szafraniec, Huy V. Vo, Camille Couprie et al.ICLR 2026
- The Computational Complexity of Circuit Discovery for Inner InterpretabilityFederico Adolfi, Martina G. Vilas, Todd WarehamICLR 2025
Builds on8
- Exact learning dynamics of deep linear networks with prior knowledgeLukas Braun, Clémentine C. J. Dominé, James Fitzgerald, Andrew M. SaxeNeurIPS 2022 · 75 citations
- Saddle-to-Saddle Dynamics in Diagonal Linear NetworksScott Pesme, Nicolas FlammarionNeurIPS 2023 · 68 citations
- The Neural Race Reduction: Dynamics of Abstraction in Gated NetworksAndrew M. Saxe, Shagun Sodhani, Sam Jay LewallenICML 2022 · 52 citations
- The Local Elasticity of Neural NetworksHangfeng He, Weijie J. SuICLR 2020 · 52 citations
- Grounding inductive biases in natural images: invariance stems from variations in dataDiane Bouchacourt, Mark Ibrahim, Ari S. MorcosNeurIPS 2021 · 32 citations
Related papers
- Feature Learning beyond the Lazy-Rich Dichotomy: Insights from Representational GeometryChi-Ning Chou, Hang Le, Yichen Wang, SueYeon ChungICML 2025
- From Lazy to Rich: Exact Learning Dynamics in Deep Linear NetworksClémentine Carla Juliette Dominé, Nicolas Anguita, Alexandra Maria Proca, Lukas Braun et al.ICLR 2025
- A self consistent theory of Gaussian Processes captures feature learning effects in finite CNNsGadi Naveh, Zohar RingelNeurIPS 2021 · 38 citations
- Mitigating the Curse of Detail: Scaling Arguments for Feature Learning and Sample ComplexityNoa Rubin, Orit Davidovich, Zohar RingelICLR 2026 · 5 citations
- Task structure and nonlinearity jointly determine learned representational geometryMatteo Alleman, Jack W. Lindsey, Stefano FusiICLR 2024 · 11 citations
