The Simplicity Bias in Multi-Task RNNs: Shared Attractors, Reuse of Dynamics, and Geometric Representation
Elia Turner, Omri Barak
摘要
How does a single interconnected neural population perform multiple tasks, each with its own dynamical requirements? The relation between task requirements and neural dynamics in Recurrent Neural Networks (RNNs) has been investigated for single tasks. The forces shaping joint dynamics of multiple tasks, however, are largely unexplored. In this work, we first construct a systematic framework to study multiple tasks in RNNs, minimizing interference from input and output correlations with the hidden representation. This allows us to reveal how RNNs tend to share attractors and reuse dynamics, a tendency we define as the "simplicity bias". We find that RNNs develop attractors sequentially during training, preferentially reusing existing dynamics and opting for simple solutions when possible. This sequenced emergence and preferential reuse encapsulate the simplicity bias. Through concrete examples, we demonstrate that new attractors primarily emerge due to task demands or architectural constraints, illustrating a balance between simplicity bias and external factors. We examine the geometry of joint representations within a single attractor, by constructing a family of tasks from a set of functions. We show that the steepness of the associated functions controls their alignment within the attractor. This arrangement again highlights the simplicity bias, as points with similar input spacings undergo comparable transformations to reach the shared attractor. Our findings propose compelling applications. The geometry of shared attractors might allow us to infer the nature of unknown tasks. Furthermore, the simplicity bias implies that without specific incentives, modularity in RNNs may not spontaneously emerge, providing insights into the conditions required for network specialization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Measuring and Controlling Solution Degeneracy across Task-Trained Recurrent Neural NetworksAnn Huang, Satpreet Harcharan Singh, Flavio Martinelli, Kanaka RajanNeurIPS 2025 · 被引用 22 次
- Discovering alternative solutions beyond the simplicity bias in recurrent neural networksWilliam Qian, Cengiz PehlevanICLR 2026 · 被引用 5 次
- Learning Dynamics of RNNs in Closed-Loop EnvironmentsYoav Ger, Omri BarakNeurIPS 2025 · 被引用 2 次
- Learning dynamics in linear recurrent neural networksAlexandra Maria Proca, Clémentine Carla Juliette Dominé, Murray Shanahan, Pedro A. M. MedianoICML 2025
- Meta-Dynamical State Space Models for Integrative Neural Data AnalysisAyesha Vermani, Josue Nassar, Hyungju Jeon, Matthew Dowling 等ICLR 2025
它引用的顶会 Paper4
- The interplay between randomness and structure during learning in RNNsFriedrich Schüßler, Francesca Mastrogiuseppe, Alexis M. Dubreuil, Srdjan Ostojic 等NeurIPS 2020 · 被引用 91 次
- Neural networks trained with SGD learn distributions of increasing complexityMaria Refinetti, Alessandro Ingrosso, Sebastian GoldtICML 2023 · 被引用 58 次
- The Neural Race Reduction: Dynamics of Abstraction in Gated NetworksAndrew M. Saxe, Shagun Sodhani, Sam Jay LewallenICML 2022 · 被引用 52 次
- Charting and Navigating the Space of Solutions for Recurrent Neural NetworksElia Turner, Kabir V. Dabholkar, Omri BarakNeurIPS 2021 · 被引用 33 次
相关 Paper
- Learning rule influences recurrent network representations but not attractor structure in decision-making tasksBrandon McMahan, Michael Kleinman, Jonathan C. KaoNeurIPS 2021 · 被引用 5 次
- Implementing Inductive bias for different navigation tasks through diverse RNN attrractorsTie Xu, Omri BarakICLR 2020 · 被引用 6 次
- Saddle-to-Saddle Dynamics Explains A Simplicity Bias Across Neural Network ArchitecturesYedi Zhang, Andrew M. Saxe, Peter E. LathamICLR 2026 · 被引用 15 次
- Shaping Sequence Attractor Schema in Recurrent Neural NetworksZhikun Chu, Bo Ho, Xiaolong Zou, Yuanyuan MiNeurIPS 2025
- A unified theory of feature learning in RNNs and DNNsJan Bauer, Kirsten Fischer, Moritz Helias, Agostina PalmigianoICML 2026 · 被引用 4 次
