Efficient Training of Minimal and Maximal Low-Rank Recurrent Neural Networks
Anushri Arora, Jonathan W. Pillow
摘要
Low-rank recurrent neural networks (RNNs) provide a powerful framework for characterizing how neural systems solve complex cognitive tasks. However, fitting and interpreting these networks remains an important open problem. In this paper, we develop new methods for efficiently fitting low-rank RNNs in "teacher-training" settings. In particular, we build upon the neural engineering framework (NEF), in which RNNs are viewed as approximating an ordinary differential equation (ODE) of interest using a set of random nonlinear basis functions. This view provides geometric insight into how the choice of neural nonlinearity (e.g. tanh, ReLU) and the distribution of model parameters affects an RNN's representational capacity. We show that this perspective leads to an online training method that achieves higher accuracy with smaller networks than previous methods such as FORCE, and outperform backprop-trained networks of similar size while requiring substantially less training time. We then consider the problem of finding minimal and maximal low-RNNs for approximating a target dynamical system. We show that a variant of orthogonal matching pursuit (OMP) can be used to find the smallest RNN for a dynamical system of interest. At the other extreme, a dual space formulation allows for efficient fitting of infinite low-rank RNNs, which provide a Gaussian Process (GP) prior over dynamical systems. We use the resulting GP marginal likelihood to optimize the hyperparameters governing neural activation functions, which leads to improved training performance even for finite RNNs. Finally, we describe active learning methods for low-rank RNNs, which speed up training through the selection of maximally informative activity patterns.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Generalized Shape Metrics on Neural RepresentationsAlex H. Williams, Erin Kunz, Simon Kornblith, Scott W. LindermanNeurIPS 2021 · 被引用 182 次
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 被引用 71 次
- Reverse-engineering recurrent neural network solutions to a hierarchical inference task for miceRylan Schaeffer, Mikail Khona, Leenoy Meshulam, International Brain Laboratory 等NeurIPS 2020 · 被引用 49 次
- Tractable Dendritic RNNs for Reconstructing Nonlinear Dynamical SystemsManuel Brenner, Florian Hess, Jonas M. Mikhaeil, Leonard F. Bereska 等ICML 2022 · 被引用 48 次
- Inferring stochastic low-rank recurrent neural networks from neural dataMatthijs Pals, A Erdem Sagtekin, Felix Pei, Manuel Glöckler 等NeurIPS 2024 · 被引用 37 次
相关 Paper
- Structured flexibility in recurrent neural networks via neuromodulationJulia Costacurta, Shaunak Bhandarkar, David M. Zoltowski, Scott W. LindermanNeurIPS 2024 · 被引用 21 次
- The interplay between randomness and structure during learning in RNNsFriedrich Schüßler, Francesca Mastrogiuseppe, Alexis M. Dubreuil, Srdjan Ostojic 等NeurIPS 2020 · 被引用 91 次
- Setting up for failure: automatic discovery of the neural mechanisms of cognitive errorsPuria Radmard, Paul M. Bays, Máté LengyelICLR 2026
- Active learning of neural population dynamics using two-photon holographic optogeneticsAndrew Wagenmaker, Lu Mi, Marton Rozsa, Matthew S. Bull 等NeurIPS 2024 · 被引用 6 次
- Low Tensor Rank Learning of Neural DynamicsArthur Pellegrino, N. Alex Cayco-Gajic, Angus ChadwickNeurIPS 2023 · 被引用 26 次
