Efficient Training of Minimal and Maximal Low-Rank Recurrent Neural Networks
Anushri Arora, Jonathan W. Pillow
Abstract
Low-rank recurrent neural networks (RNNs) provide a powerful framework for characterizing how neural systems solve complex cognitive tasks. However, fitting and interpreting these networks remains an important open problem. In this paper, we develop new methods for efficiently fitting low-rank RNNs in "teacher-training" settings. In particular, we build upon the neural engineering framework (NEF), in which RNNs are viewed as approximating an ordinary differential equation (ODE) of interest using a set of random nonlinear basis functions. This view provides geometric insight into how the choice of neural nonlinearity (e.g. tanh, ReLU) and the distribution of model parameters affects an RNN's representational capacity. We show that this perspective leads to an online training method that achieves higher accuracy with smaller networks than previous methods such as FORCE, and outperform backprop-trained networks of similar size while requiring substantially less training time. We then consider the problem of finding minimal and maximal low-RNNs for approximating a target dynamical system. We show that a variant of orthogonal matching pursuit (OMP) can be used to find the smallest RNN for a dynamical system of interest. At the other extreme, a dual space formulation allows for efficient fitting of infinite low-rank RNNs, which provide a Gaussian Process (GP) prior over dynamical systems. We use the resulting GP marginal likelihood to optimize the hyperparameters governing neural activation functions, which leads to improved training performance even for finite RNNs. Finally, we describe active learning methods for low-rank RNNs, which speed up training through the selection of maximally informative activity patterns.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e7855b8-e6bd-4161-816e-b31e1d4eaaf5Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Generalized Shape Metrics on Neural RepresentationsAlex H. Williams, Erin Kunz, Simon Kornblith, Scott W. LindermanNeurIPS 2021 · 182 citations
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 71 citations
- Reverse-engineering recurrent neural network solutions to a hierarchical inference task for miceRylan Schaeffer, Mikail Khona, Leenoy Meshulam, International Brain Laboratory et al.NeurIPS 2020 · 49 citations
- Tractable Dendritic RNNs for Reconstructing Nonlinear Dynamical SystemsManuel Brenner, Florian Hess, Jonas M. Mikhaeil, Leonard F. Bereska et al.ICML 2022 · 48 citations
- Inferring stochastic low-rank recurrent neural networks from neural dataMatthijs Pals, A Erdem Sagtekin, Felix Pei, Manuel Glöckler et al.NeurIPS 2024 · 37 citations
Related papers
- Structured flexibility in recurrent neural networks via neuromodulationJulia Costacurta, Shaunak Bhandarkar, David M. Zoltowski, Scott W. LindermanNeurIPS 2024 · 21 citations
- The interplay between randomness and structure during learning in RNNsFriedrich Schüßler, Francesca Mastrogiuseppe, Alexis M. Dubreuil, Srdjan Ostojic et al.NeurIPS 2020 · 91 citations
- Setting up for failure: automatic discovery of the neural mechanisms of cognitive errorsPuria Radmard, Paul M. Bays, Máté LengyelICLR 2026
- Active learning of neural population dynamics using two-photon holographic optogeneticsAndrew Wagenmaker, Lu Mi, Marton Rozsa, Matthew S. Bull et al.NeurIPS 2024 · 6 citations
- Low Tensor Rank Learning of Neural DynamicsArthur Pellegrino, N. Alex Cayco-Gajic, Angus ChadwickNeurIPS 2023 · 26 citations
