Training biologically plausible recurrent neural networks on cognitive tasks with long-term dependencies
Wayne Soo, Vishwa Goudar, Xiao-Jing Wang
摘要
Training recurrent neural networks (RNNs) has become a go-to approach for generating and evaluating mechanistic neural hypotheses for cognition. The ease and efficiency of training RNNs with backpropagation through time and the availability of robustly supported deep learning libraries has made RNN modeling more approachable and accessible to neuroscience. Yet, a major technical hindrance remains. Cognitive processes such as working memory and decision making involve neural population dynamics over a long period of time within a behavioral trial and across trials. It is difficult to train RNNs to accomplish tasks where neural representations and dynamics have long temporal dependencies without gating mechanisms such as LSTMs or GRUs which currently lack experimental support and prohibit direct comparison between RNNs and biological neural circuits. We tackled this problem based on the idea of specialized skip-connections through time to support the emergence of task-relevant dynamics, and subsequently reinstitute biological plausibility by reverting to the original architecture. We show that this approach enables RNNs to successfully learn cognitive tasks that prove impractical if not impossible to learn using conventional methods. Over numerous tasks considered here, we achieve less training steps and shorter wall-clock times, particularly in tasks that require learning long-term dependencies via temporal integration over long timescales or maintaining a memory of past events in hidden-states. Our methods expand the range of experimental tasks that biologically plausible RNN models can learn, thereby supporting the development of theory for the emergent neural mechanisms of computations involving long-term dependencies.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Unconditional stability of a recurrent neural circuit implementing divisive normalizationShivang Rawat, David J. Heeger, Stefano MartinianiNeurIPS 2024 · 被引用 8 次
- Recurrent neural network dynamical systems for biological visionWayne Soo, Aldo Battista, Puria Radmard, Xiao-Jing WangNeurIPS 2024 · 被引用 7 次
- Second-order forward-mode optimization of recurrent neural networks for neuroscienceYoujing Yu, Rui Xia, Qingxi Ma, Máté Lengyel 等NeurIPS 2024 · 被引用 6 次
- Disentangling Representations through Multi-task LearningPantelis Vafidis, Aman Bhargava, Antonio RangelICLR 2025
它引用的顶会 Paper5
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer 等NeurIPS 2021 · 被引用 3,862 次
- The interplay between randomness and structure during learning in RNNsFriedrich Schüßler, Francesca Mastrogiuseppe, Alexis M. Dubreuil, Srdjan Ostojic 等NeurIPS 2020 · 被引用 91 次
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 被引用 71 次
- A mechanistic multi-area recurrent network model of decision-makingMichael Kleinman, Chandramouli Chandrasekaran, Jonathan C. KaoNeurIPS 2021 · 被引用 19 次
- Training stochastic stabilized supralinear networks by dynamics-neutral growthWayne Soo, Máté LengyelNeurIPS 2022 · 被引用 7 次
相关 Paper
- RNNs Incrementally Evolving on an Equilibrium Manifold: A Panacea for Vanishing and Exploding Gradients?Anil Kag, Ziming Zhang, Venkatesh SaligramaICLR 2020 · 被引用 51 次
- Short-Term Plasticity Neurons Learning to Learn and ForgetHector Garcia Rodriguez, Qinghai Guo, Timoleon MoraitisICML 2022 · 被引用 15 次
- Setting up for failure: automatic discovery of the neural mechanisms of cognitive errorsPuria Radmard, Paul M. Bays, Máté LengyelICLR 2026
- Skipper: Enabling efficient SNN training through activation-checkpointing and time-skippingSonali Singh, Anup Sarma, Sen Lu, Abhronil Sengupta 等MICRO 2022 · 被引用 13 次
- Emergent mechanisms for long timescales depend on training curriculum and affect performance in memory tasksSina Khajehabdollahi, Roxana Zeraati, Emmanouil Giannakakis, Tim Jakob Schäfer 等ICLR 2024 · 被引用 8 次
