Hidden Traveling Waves bind Working Memory Variables in Recurrent Neural Networks
Arjun Karuvally, Terrence J. Sejnowski, Hava T. Siegelmann
Abstract
Traveling waves are a fundamental phenomenon in the brain, playing a crucial role in short-term information storage. In this study, we leverage the concept of traveling wave dynamics within a neural lattice to formulate a theoretical model of neural working memory, study its properties, and its real world implications in AI. The proposed model diverges from traditional approaches, which assume information storage in static, register-like locations updated by interference. Instead, the model stores data as waves that is updated by the wave's boundary conditions. We rigorously examine the model's capabilities in representing and learning state histories, which are vital for learning history-dependent dynamical systems. The findings reveal that the model reliably stores external information and enhances the learning process by addressing the diminishing gradient problem. To understand the model's real-world applicability, we explore two cases: linear boundary condition (LBC) and non-linear, self-attention-driven boundary condition (SBC). The model with the linear boundary condition results in a shift matrix plus low-rank matrix currently used in H3 state space RNN. Further, our experiments with LBC reveal that this matrix is effectively learned by Recurrent Neural Networks (RNNs) through backpropagation when modeling history-dependent dynamical systems. Conversely, the SBC parallels the autoregressive loop of an attention-only transformer with the context vector representing the wave substrate. Collectively, our findings suggest the broader relevance of traveling waves in AI and its potential in advancing neural network architectures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Bridging Expressivity and Scalability with Adaptive Unitary SSMsArjun Karuvally, Franz Nowak, T. Anderson Keller, Carmen Amo Alonso et al.NeurIPS 2025 · 8 citations
- Optimizing Neural Network Representations of Boolean NetworksJoshua Russell, Ignacio Gavier, Devdhar Patel, Edward A. Rietman et al.ICLR 2025
Builds on3
- Hungry Hungry Hippos: Towards Language Modeling with State Space ModelsDaniel Y. Fu, Tri Dao, Khaled Kamal Saab, Armin W. Thomas et al.ICLR 2023 · 117 citations
- Traveling Waves Encode The Recent Past and Enhance Sequence LearningT. Anderson Keller, Lyle Muller, Terrence J. Sejnowski, Max WellingICLR 2024 · 26 citations
- How to Train your HIPPO: State Space Models with Generalized Orthogonal Basis ProjectionsAlbert Gu, Isys Johnson, Aman Timalsina, Atri Rudra et al.ICLR 2023 · 11 citations
Related papers
- Short-Term Plasticity Neurons Learning to Learn and ForgetHector Garcia Rodriguez, Qinghai Guo, Timoleon MoraitisICML 2022 · 15 citations
- Neural Wave Machines: Learning Spatiotemporally Structured Representations with Locally Coupled Oscillatory Recurrent Neural NetworksT. Anderson Keller, Max WellingICML 2023 · 26 citations
- On the difficulty of learning chaotic dynamics with RNNsJonas M. Mikhaeil, Zahra Monfared, Daniel DurstewitzNeurIPS 2022 · 109 citations
- Coupled Oscillatory Recurrent Neural Network (coRNN): An accurate and (gradient) stable architecture for learning long time dependenciesT. Konstantin Rusch, Siddhartha MishraICLR 2021 · 121 citations
- Dynamics and Representation Structure of Local Approximations to Gradient-Based Learning in Linear Recurrent Neural NetworksEzekiel Williams, Alexandre Payeur, Guillaume LajoieICML 2026
