Information dynamics and Memory in Neural Networks through Fisher Information Diffusion
Haodong Qin, Tatyana Sharpee
Abstract
We present a general theoretical framework for analyzing how information about past inputs is encoded in recurrent networks into evolving dynamics rather than being represented as convergence to static attractors. Using dynamic mean-field theory and diffusion from physics, we derive a Fisher information diffusion operator that links network connectivity structure to the time-resolved propagation of information across interacting subpopulations. The analysis reveals that operating near criticality (spectral radius near one) is necessary but not sufficient for reliable memory in structured or non-normal recurrent networks; effective information retention requires alignment between input–output structure and stable dynamical subspaces. The theory yields principled initialization rules that balance stability and sensitivity, mitigating vanishing and exploding gradients. Experiments on the copy task and sequential MNIST show faster convergence and higher accuracy than standard random initialization. Together, these results provide both principled design guidelines for recurrent networks and new theoretical insight into how information can be preserved over time in their dynamics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 01e2915d-acb5-42c5-9a4c-8fe1bfdc1da6Builds on6
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 3,482 citations
- Hopfield Networks is All You NeedHubert Ramsauer, Bernhard Schäfl, Johannes Lehner, Philipp Seidl et al.ICLR 2021 · 620 citations
- Understanding Approximate Fisher Information for Fast Convergence of Natural Gradient Descent in Wide Neural NetworksRyo Karakida, Kazuki OsawaNeurIPS 2020 · 39 citations
- Traveling Waves Encode The Recent Past and Enhance Sequence LearningT. Anderson Keller, Lyle Muller, Terrence J. Sejnowski, Max WellingICLR 2024 · 26 citations
- Improved memory in recurrent neural networks with sequential non-normal dynamicsA. Emin Orhan, Xaq PitkowICLR 2020 · 16 citations
Related papers
- Information Geometry of Orthogonal Initializations and TrainingPiotr Aleksander Sokól, Il Memming ParkICLR 2020 · 17 citations
- A time-resolved theory of information encoding in recurrent neural networksRainer Engelken, Sven GoedekeNeurIPS 2022 · 3 citations
- Revisiting Glorot Initialization for Long-Range Linear RecurrencesNoga Bar, Mariia Seleznova, Yotam Alexander, Gitta Kutyniok et al.NeurIPS 2025 · 3 citations
- How connectivity structure shapes rich and lazy learning in neural circuitsYuhan Helena Liu, Aristide Baratin, Jonathan Cornford, Stefan Mihalas et al.ICLR 2024 · 26 citations
- Operative dimensions in unconstrained connectivity of recurrent neural networksRenate Krause, Matthew Cook, Sepp Kollmorgen, Valerio Mante et al.NeurIPS 2022 · 12 citations
