Information dynamics and Memory in Neural Networks through Fisher Information Diffusion
Haodong Qin, Tatyana Sharpee
摘要
We present a general theoretical framework for analyzing how information about past inputs is encoded in recurrent networks into evolving dynamics rather than being represented as convergence to static attractors. Using dynamic mean-field theory and diffusion from physics, we derive a Fisher information diffusion operator that links network connectivity structure to the time-resolved propagation of information across interacting subpopulations. The analysis reveals that operating near criticality (spectral radius near one) is necessary but not sufficient for reliable memory in structured or non-normal recurrent networks; effective information retention requires alignment between input–output structure and stable dynamical subspaces. The theory yields principled initialization rules that balance stability and sensitivity, mitigating vanishing and exploding gradients. Experiments on the copy task and sequential MNIST show faster convergence and higher accuracy than standard random initialization. Together, these results provide both principled design guidelines for recurrent networks and new theoretical insight into how information can be preserved over time in their dynamics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Hopfield Networks is All You NeedHubert Ramsauer, Bernhard Schäfl, Johannes Lehner, Philipp Seidl 等ICLR 2021 · 被引用 620 次
- Understanding Approximate Fisher Information for Fast Convergence of Natural Gradient Descent in Wide Neural NetworksRyo Karakida, Kazuki OsawaNeurIPS 2020 · 被引用 39 次
- Traveling Waves Encode The Recent Past and Enhance Sequence LearningT. Anderson Keller, Lyle Muller, Terrence J. Sejnowski, Max WellingICLR 2024 · 被引用 26 次
- Improved memory in recurrent neural networks with sequential non-normal dynamicsA. Emin Orhan, Xaq PitkowICLR 2020 · 被引用 16 次
相关 Paper
- Information Geometry of Orthogonal Initializations and TrainingPiotr Aleksander Sokól, Il Memming ParkICLR 2020 · 被引用 17 次
- A time-resolved theory of information encoding in recurrent neural networksRainer Engelken, Sven GoedekeNeurIPS 2022 · 被引用 3 次
- Revisiting Glorot Initialization for Long-Range Linear RecurrencesNoga Bar, Mariia Seleznova, Yotam Alexander, Gitta Kutyniok 等NeurIPS 2025 · 被引用 3 次
- How connectivity structure shapes rich and lazy learning in neural circuitsYuhan Helena Liu, Aristide Baratin, Jonathan Cornford, Stefan Mihalas 等ICLR 2024 · 被引用 26 次
- Operative dimensions in unconstrained connectivity of recurrent neural networksRenate Krause, Matthew Cook, Sepp Kollmorgen, Valerio Mante 等NeurIPS 2022 · 被引用 12 次
