Discovering alternative solutions beyond the simplicity bias in recurrent neural networks
William Qian, Cengiz Pehlevan
摘要
Training recurrent neural networks (RNNs) to perform neuroscience-style tasks has become a popular way to generate hypotheses for how neural circuits in the brain might perform computations. Recent work has demonstrated that task-trained RNNs possess a strong simplicity bias. In particular, this inductive bias often causes RNNs trained on the same task to collapse on effectively the same solution, typically comprised of fixed-point attractors or other low-dimensional dynamical motifs. While such solutions are readily interpretable, this collapse proves counterproductive for the sake of generating a set of genuinely unique hypotheses for how neural computations might be performed. Here we propose Iterative Neural Similarity Deflation (INSD), a simple method to break this inductive bias. By penalizing linear predictivity of neural activity produced by standard task-trained RNNs, we find an alternative class of solutions to classic neuroscience-style RNN tasks. These solutions appear distinct across a battery of analysis techniques, including representational similarity metrics, dynamical systems analysis, and the linear decodability of task-relevant variables. Moreover, these alternative solutions can sometimes achieve superior performance in difficult or out-of-distribution task regimes. Our findings underscore the importance of moving beyond the simplicity bias to uncover richer and more varied models of neural computation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper16
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Generalized Shape Metrics on Neural RepresentationsAlex H. Williams, Erin Kunz, Simon Kornblith, Scott W. LindermanNeurIPS 2021 · 被引用 182 次
- The interplay between randomness and structure during learning in RNNsFriedrich Schüßler, Francesca Mastrogiuseppe, Alexis M. Dubreuil, Srdjan Ostojic 等NeurIPS 2020 · 被引用 91 次
- Linear Adversarial Concept ErasureShauli Ravfogel, Michael Twiton, Yoav Goldberg, Ryan CotterellICML 2022 · 被引用 89 次
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 被引用 71 次
相关 Paper
- Measuring and Controlling Solution Degeneracy across Task-Trained Recurrent Neural NetworksAnn Huang, Satpreet Harcharan Singh, Flavio Martinelli, Kanaka RajanNeurIPS 2025 · 被引用 22 次
- The Simplicity Bias in Multi-Task RNNs: Shared Attractors, Reuse of Dynamics, and Geometric RepresentationElia Turner, Omri BarakNeurIPS 2023 · 被引用 25 次
- Learning rule influences recurrent network representations but not attractor structure in decision-making tasksBrandon McMahan, Michael Kleinman, Jonathan C. KaoNeurIPS 2021 · 被引用 5 次
- Reverse-engineering recurrent neural network solutions to a hierarchical inference task for miceRylan Schaeffer, Mikail Khona, Leenoy Meshulam, International Brain Laboratory 等NeurIPS 2020 · 被引用 49 次
- Charting and Navigating the Space of Solutions for Recurrent Neural NetworksElia Turner, Kabir V. Dabholkar, Omri BarakNeurIPS 2021 · 被引用 33 次
