Reverse engineering recurrent neural networks with Jacobian switching linear dynamical systems
Jimmy T. H. Smith, Scott W. Linderman, David Sussillo
摘要
Recurrent neural networks (RNNs) are powerful models for processing time-series data, but it remains challenging to understand how they function. Improving this understanding is of substantial interest to both the machine learning and neuroscience communities. The framework of reverse engineering a trained RNN by linearizing around its fixed points has provided insight, but the approach has significant challenges. These include difficulty choosing which fixed point to expand around when studying RNN dynamics and error accumulation when reconstructing the nonlinear dynamics with the linearized dynamics. We present a new model that overcomes these limitations by co-training an RNN with a novel switching linear dynamical system (SLDS) formulation. A first-order Taylor series expansion of the co-trained RNN and an auxiliary function trained to pick out the RNN's fixed points govern the SLDS dynamics. The results are a trained SLDS variant that closely approximates the RNN, an auxiliary function that can produce a fixed point for each point in state-space, and a trained nonlinear RNN whose dynamics have been regularized such that its first-order terms perform the computation, if possible. This model removes the post-training fixed point optimization and allows us to unambiguously study the learned dynamics of the SLDS at any point in state-space. It also generalizes SLDS models to continuous manifolds of switching points while sharing parameters across switches. We validate the utility of the model on two synthetic tasks relevant to previous work reverse engineering RNNs. We then show that our model can be used as a drop-in in more complex architectures, such as LFADS, and apply this LFADS hybrid to analyze single-trial spiking activity from the motor system of a non-human primate.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 被引用 71 次
- Generalized Teacher Forcing for Learning Chaotic DynamicsFlorian Hess, Zahra Monfared, Manuel Brenner, Daniel DurstewitzICML 2023 · 被引用 67 次
- Towards Scalable and Stable Parallelization of Nonlinear RNNsXavier Gonzalez, Andrew Warrington, Jimmy T. H. Smith, Scott W. LindermanNeurIPS 2024 · 被引用 47 次
- Bifurcations and loss jumps in RNN trainingLukas Eisenmann, Zahra Monfared, Niclas Alexander Göring, Daniel DurstewitzNeurIPS 2023 · 被引用 29 次
- Modeling Latent Neural Dynamics with Gaussian Process Switching Linear Dynamical SystemsAmber Hu, David M. Zoltowski, Aditya Nair, David Anderson 等NeurIPS 2024 · 被引用 22 次
它引用的顶会 Paper5
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- Recurrent Switching Dynamical Systems Models for Multiple Interacting Neural PopulationsJoshua I. Glaser, Matthew R. Whiteway, John P. Cunningham, Liam Paninski 等NeurIPS 2020 · 被引用 113 次
- Reverse-engineering recurrent neural network solutions to a hierarchical inference task for miceRylan Schaeffer, Mikail Khona, Leenoy Meshulam, International Brain Laboratory 等NeurIPS 2020 · 被引用 49 次
- Identifying nonlinear dynamical systems with multiple time scales and long-range dependenciesDominik Schmidt, Georgia Koppe, Zahra Monfared, Max Beutelspacher 等ICLR 2021 · 被引用 41 次
- How recurrent networks implement contextual processing in sentiment analysisNiru Maheswaranathan, David SussilloICML 2020 · 被引用 25 次
相关 Paper
- Inference of Neural Dynamics Using Switching Recurrent Neural NetworksYongxu Zhang, Shreya SaxenaNeurIPS 2024 · 被引用 8 次
- Detecting Invariant Manifolds in ReLU-Based RNNsLukas Eisenmann, Alena Brändle, Zahra Monfared, Daniel DurstewitzICLR 2026 · 被引用 3 次
- Almost-Linear RNNs Yield Highly Interpretable Symbolic Codes in Dynamical Systems ReconstructionManuel Brenner, Christoph Jürgen Hemmer, Zahra Monfared, Daniel DurstewitzNeurIPS 2024 · 被引用 22 次
- Parsing neural dynamics with infinite recurrent switching linear dynamical systemsVictor Geadah, International Brain Laboratory, Jonathan W. PillowICLR 2024 · 被引用 7 次
- Inferring stochastic low-rank recurrent neural networks from neural dataMatthijs Pals, A Erdem Sagtekin, Felix Pei, Manuel Glöckler 等NeurIPS 2024 · 被引用 37 次
