Training biologically plausible recurrent neural networks on cognitive tasks with long-term dependencies
Wayne Soo, Vishwa Goudar, Xiao-Jing Wang
Abstract
Training recurrent neural networks (RNNs) has become a go-to approach for generating and evaluating mechanistic neural hypotheses for cognition. The ease and efficiency of training RNNs with backpropagation through time and the availability of robustly supported deep learning libraries has made RNN modeling more approachable and accessible to neuroscience. Yet, a major technical hindrance remains. Cognitive processes such as working memory and decision making involve neural population dynamics over a long period of time within a behavioral trial and across trials. It is difficult to train RNNs to accomplish tasks where neural representations and dynamics have long temporal dependencies without gating mechanisms such as LSTMs or GRUs which currently lack experimental support and prohibit direct comparison between RNNs and biological neural circuits. We tackled this problem based on the idea of specialized skip-connections through time to support the emergence of task-relevant dynamics, and subsequently reinstitute biological plausibility by reverting to the original architecture. We show that this approach enables RNNs to successfully learn cognitive tasks that prove impractical if not impossible to learn using conventional methods. Over numerous tasks considered here, we achieve less training steps and shorter wall-clock times, particularly in tasks that require learning long-term dependencies via temporal integration over long timescales or maintaining a memory of past events in hidden-states. Our methods expand the range of experimental tasks that biologically plausible RNN models can learn, thereby supporting the development of theory for the emergent neural mechanisms of computations involving long-term dependencies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Unconditional stability of a recurrent neural circuit implementing divisive normalizationShivang Rawat, David J. Heeger, Stefano MartinianiNeurIPS 2024 · 8 citations
- Recurrent neural network dynamical systems for biological visionWayne Soo, Aldo Battista, Puria Radmard, Xiao-Jing WangNeurIPS 2024 · 7 citations
- Second-order forward-mode optimization of recurrent neural networks for neuroscienceYoujing Yu, Rui Xia, Qingxi Ma, Máté Lengyel et al.NeurIPS 2024 · 6 citations
- Disentangling Representations through Multi-task LearningPantelis Vafidis, Aman Bhargava, Antonio RangelICLR 2025
Builds on5
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer et al.NeurIPS 2021 · 3,862 citations
- The interplay between randomness and structure during learning in RNNsFriedrich Schüßler, Francesca Mastrogiuseppe, Alexis M. Dubreuil, Srdjan Ostojic et al.NeurIPS 2020 · 91 citations
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 71 citations
- A mechanistic multi-area recurrent network model of decision-makingMichael Kleinman, Chandramouli Chandrasekaran, Jonathan C. KaoNeurIPS 2021 · 19 citations
- Training stochastic stabilized supralinear networks by dynamics-neutral growthWayne Soo, Máté LengyelNeurIPS 2022 · 7 citations
Related papers
- RNNs Incrementally Evolving on an Equilibrium Manifold: A Panacea for Vanishing and Exploding Gradients?Anil Kag, Ziming Zhang, Venkatesh SaligramaICLR 2020 · 51 citations
- Short-Term Plasticity Neurons Learning to Learn and ForgetHector Garcia Rodriguez, Qinghai Guo, Timoleon MoraitisICML 2022 · 15 citations
- Setting up for failure: automatic discovery of the neural mechanisms of cognitive errorsPuria Radmard, Paul M. Bays, Máté LengyelICLR 2026
- Skipper: Enabling efficient SNN training through activation-checkpointing and time-skippingSonali Singh, Anup Sarma, Sen Lu, Abhronil Sengupta et al.MICRO 2022 · 13 citations
- Emergent mechanisms for long timescales depend on training curriculum and affect performance in memory tasksSina Khajehabdollahi, Roxana Zeraati, Emmanouil Giannakakis, Tim Jakob Schäfer et al.ICLR 2024 · 8 citations
