Reverse-engineering recurrent neural network solutions to a hierarchical inference task for mice
Rylan Schaeffer, Mikail Khona, Leenoy Meshulam, International Brain Laboratory, Ila Fiete
Abstract
We study how recurrent neural networks (RNNs) solve a hierarchical inference task involving two latent variables and disparate timescales separated by 1-2 orders of magnitude. The task is of interest to the International Brain Laboratory, a global collaboration of experimental and theoretical neuroscientists studying how the mammalian brain generates behavior. We make four discoveries. First, RNNs learn behavior that is quantitatively similar to ideal Bayesian baselines. Second, RNNs perform inference by learning a two-dimensional subspace defining beliefs about the latent variables. Third, the geometry of RNN dynamics reflects an induced coupling between the two separate inference processes necessary to solve the task. Fourth, we perform model compression through a novel form of knowledge distillation on hidden representations – Representations and Dynamics Distillation (RADD)– to reduce the RNN dynamics to a low-dimensional, highly interpretable model. This technique promises a useful tool for interpretability of high dimensional nonlinear dynamical systems. Altogether, this work yields predictions to guide exploration and analysis of mouse neural data and circuity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dbbdd032-7da6-432d-9092-7c1245cd2308Cited by top-tier papers10
- No Free Lunch from Deep Learning in Neuroscience: A Case Study through Models of the Entorhinal-Hippocampal CircuitRylan Schaeffer, Mikail Khona, Ila FieteNeurIPS 2022 · 81 citations
- Extracting computational mechanisms from neural data using low-rank RNNsAdrian Valente, Jonathan W. Pillow, Srdjan OstojicNeurIPS 2022 · 71 citations
- Beyond Geometry: Comparing the Temporal Structure of Computation in Neural Circuits with Dynamical Similarity AnalysisMitchell Ostrow, Adam Eisen, Leo Kozachkov, Ila FieteNeurIPS 2023 · 60 citations
- Reverse engineering recurrent neural networks with Jacobian switching linear dynamical systemsJimmy T. H. Smith, Scott W. Linderman, David SussilloNeurIPS 2021 · 44 citations
- Self-Supervised Learning of Representations for Space Generates Multi-Modular Grid CellsRylan Schaeffer, Mikail Khona, Tzuhsuan Ma, Cristóbal Eyzaguirre et al.NeurIPS 2023 · 40 citations
Related papers
- Discovering alternative solutions beyond the simplicity bias in recurrent neural networksWilliam Qian, Cengiz PehlevanICLR 2026 · 5 citations
- Charting and Navigating the Space of Solutions for Recurrent Neural NetworksElia Turner, Kabir V. Dabholkar, Omri BarakNeurIPS 2021 · 33 citations
- Reconstructing Nonlinear Dynamical Systems from Multi-Modal Time SeriesDaniel Kramer, Philine Lou Bommer, Daniel Durstewitz, Carlo Tombolini et al.ICML 2022 · 25 citations
- Flow-field inference from neural data using deep recurrent networksTimothy Doyeon Kim, Thomas Zhihao Luo, Tankut Can, Kamesh Krishnamurthy et al.ICML 2025
- High-dimensional neuronal activity from low-dimensional latent dynamics: a solvable modelValentin Schmutz, Ali Haydaroglu, Shuqi Wang, Yixiao Feng et al.NeurIPS 2025 · 9 citations
