Dynamical phases of short-term memory mechanisms in RNNs
Bariscan Kurtkaya, Fatih Dinc, Mert Yüksekgönül, Marta Blanco-Pozo, Ege Çirakman, Mark J. Schnitzer, Yucel Yemez, Hidenori Tanaka, Peng Yuan, Nina Miolane
摘要
Short-term memory is essential for cognitive processing, yet our understanding of its neural mechanisms remains unclear. Neuroscience has long focused on how sequential activity patterns, where neurons fire one after another within large networks, can explain how information is maintained. While recurrent connections were shown to drive sequential dynamics, a mechanistic understanding of this process still remains unknown. In this work, we introduce two unique mechanisms that can support this form of short-term memory: slowpoint manifolds generating direct sequences or limit cycles providing temporally localized approximations. Using analytical models, we identify fundamental properties that govern the selection of each mechanism. Precisely, on shortterm memory tasks (delayed cue-discrimination tasks), we derive theoretical scaling laws for critical learning rates as a function of the delay period length, beyond which no learning is possible. We empirically verify these results by training and evaluating approximately 80,000 recurrent neural networks (RNNs), which are publicly available for further analysis 1 . Overall, our work provides new insights into short-term memory mechanisms and proposes experimentally testable predictions for systems neuroscience.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Measuring and Controlling Solution Degeneracy across Task-Trained Recurrent Neural NetworksAnn Huang, Satpreet Harcharan Singh, Flavio Martinelli, Kanaka RajanNeurIPS 2025 · 被引用 22 次
- Discovering alternative solutions beyond the simplicity bias in recurrent neural networksWilliam Qian, Cengiz PehlevanICLR 2026 · 被引用 5 次
- Information dynamics and Memory in Neural Networks through Fisher Information DiffusionHaodong Qin, Tatyana SharpeeICML 2026
它引用的顶会 Paper5
- Pretraining task diversity and the emergence of non-Bayesian in-context learning for regressionAllan Raventós, Mansheej Paul, Feng Chen, Surya GanguliNeurIPS 2023 · 被引用 174 次
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 被引用 71 次
- Progress measures for grokking via mechanistic interpretabilityNeel Nanda, Lawrence Chan, Tom Lieberum, Jess Smith 等ICLR 2023 · 被引用 54 次
- Competition Dynamics Shape Algorithmic Phases of In-Context LearningCore Francisco Park, Ekdeep Singh Lubana, Hidenori TanakaICLR 2025
- A Percolation Model of Emergence: Analyzing Transformers Trained on a Formal LanguageEkdeep Singh Lubana, Kyogo Kawaguchi, Robert P. Dick, Hidenori TanakaICLR 2025
相关 Paper
- Training biologically plausible recurrent neural networks on cognitive tasks with long-term dependenciesWayne Soo, Vishwa Goudar, Xiao-Jing WangNeurIPS 2023 · 被引用 16 次
- Back to the Continuous AttractorÁbel Ságodi, Guillermo Martín-Sánchez, Piotr A. Sokól, Il Memming ParkNeurIPS 2024 · 被引用 21 次
- Temporal superposition and feature geometry of RNNs under memory demandsPratyaksh Sharma, Alexandra Maria Proca, Lucas Prieto, Pedro A. M. MedianoICLR 2026
- Charting and Navigating the Space of Solutions for Recurrent Neural NetworksElia Turner, Kabir V. Dabholkar, Omri BarakNeurIPS 2021 · 被引用 33 次
- Setting up for failure: automatic discovery of the neural mechanisms of cognitive errorsPuria Radmard, Paul M. Bays, Máté LengyelICLR 2026
