Mechanistic Interpretability of RNNs emulating Hidden Markov Models
Elia Torre, Michele Viscione, Lucas Pompe, Benjamin F. Grewe, Valerio Mante
摘要
Recurrent neural networks (RNNs) provide a powerful approach in neuroscience to infer latent dynamics in neural populations and to generate hypotheses about the neural computations underlying behavior. However, past work has focused on relatively simple, input-driven, and largely deterministic behaviors - little is known about the mechanisms that would allow RNNs to generate the richer, spontaneous, and potentially stochastic behaviors observed in natural settings. Modeling with Hidden Markov Models (HMMs) has revealed a segmentation of natural behaviors into discrete latent states with stochastic transitions between them, a type of dynamics that may appear at odds with the continuous state spaces implemented by RNNs. Here we first show that RNNs can replicate HMM emission statistics and then reverse-engineer the trained networks to uncover the mechanisms they implement. In the absence of inputs, the activity of trained RNNs collapses towards a single fixed point. When driven by stochastic input, trajectories instead exhibit noise-sustained dynamics along closed orbits. Rotation along these orbits modulates the emission probabilities and is governed by transitions between regions of slow, noise-driven dynamics connected by fast, deterministic transitions. The trained RNNs develop highly structured connectivity, with a small set of"kick neurons"initiating transitions between these regions. This mechanism emerges during training as the network shifts into a regime of stochastic resonance, enabling it to perform probabilistic computations. Analyses across multiple HMM architectures - fully connected, cyclic, and linear-chain - reveal that this solution generalizes through the modular reuse of the same dynamical motif, suggesting a compositional principle by which RNNs can emulate complex discrete latent dynamics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Reverse engineering recurrent neural networks with Jacobian switching linear dynamical systemsJimmy T. H. Smith, Scott W. Linderman, David SussilloNeurIPS 2021 · 被引用 44 次
- Operative dimensions in unconstrained connectivity of recurrent neural networksRenate Krause, Matthew Cook, Sepp Kollmorgen, Valerio Mante 等NeurIPS 2022 · 被引用 12 次
- Trial matching: capturing variability with data-constrained spiking neural networksChristos Sourmpis, Carl C. H. Petersen, Wulfram Gerstner, Guillaume BellecNeurIPS 2023 · 被引用 9 次
相关 Paper
- Inference of Neural Dynamics Using Switching Recurrent Neural NetworksYongxu Zhang, Shreya SaxenaNeurIPS 2024 · 被引用 8 次
- Inferring stochastic low-rank recurrent neural networks from neural dataMatthijs Pals, A Erdem Sagtekin, Felix Pei, Manuel Glöckler 等NeurIPS 2024 · 被引用 37 次
- Identifying Connectivity Distributions from Neural Dynamics Using FlowsTimothy Kim, Ulises Obilinovic, Yiliu Wang, Eric SheaBrown 等ICML 2026 · 被引用 1 次
- On Scrambling Phenomena for Randomly Initialized Recurrent NetworksVaggos Chatziafratis, Ioannis Panageas, Clayton Sanford, Stelios StavroulakisNeurIPS 2022 · 被引用 3 次
- Noisy Recurrent Neural NetworksSoon Hoe Lim, N. Benjamin Erichson, Liam Hodgkinson, Michael W. MahoneyNeurIPS 2021 · 被引用 77 次
