DeepSITH: Efficient Learning via Decomposition of What and When Across Time Scales
Brandon G. Jacques, Zoran Tiganj, Marc W. Howard, Per B. Sederberg
Abstract
Extracting temporal relationships over a range of scales is a hallmark of human perception and cognition-and thus it is a critical feature of machine learning applied to real-world problems. Neural networks are either plagued by the exploding/vanishing gradient problem in recurrent neural networks (RNNs) or must adjust their parameters to learn the relevant time scales (e.g., in LSTMs). This paper introduces DeepSITH, a deep network comprising biologically-inspired Scale-Invariant Temporal History (SITH) modules in series with dense connections between layers. Each SITH module is simply a set of time cells coding what happened when with a geometrically-spaced set of time lags. The dense connections between layers change the definition of what from one layer to the next. The geometric series of time lags implies that the network codes time on a logarithmic scale, enabling DeepSITH network to learn problems requiring memory over a wide range of time scales. We compare DeepSITH to LSTMs and other recent RNNs on several time series prediction and decoding tasks. DeepSITH achieves results comparable to state-of-the-art performance on these problems and continues to perform well even as the delays are increased.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9339099a-d945-4334-8c64-5d7dcd7ad8ebCited by top-tier papers2
- Traveling Waves Encode The Recent Past and Enhance Sequence LearningT. Anderson Keller, Lyle Muller, Terrence J. Sejnowski, Max WellingICLR 2024 · 26 citations
- Deep Reinforcement Learning with Time-Scale Invariant MemoryMd Rysul Kabir, James Mochizuki-Freeman, Zoran TiganjAAAI 2025
Builds on1
Related papers
- A deep convolutional neural network that is invariant to time rescalingBrandon G. Jacques, Zoran Tiganj, Aakash Sarkar, Marc W. Howard et al.ICML 2022 · 10 citations
- Training biologically plausible recurrent neural networks on cognitive tasks with long-term dependenciesWayne Soo, Vishwa Goudar, Xiao-Jing WangNeurIPS 2023 · 16 citations
- Identifying nonlinear dynamical systems with multiple time scales and long-range dependenciesDominik Schmidt, Georgia Koppe, Zahra Monfared, Max Beutelspacher et al.ICLR 2021 · 41 citations
- Emergent mechanisms for long timescales depend on training curriculum and affect performance in memory tasksSina Khajehabdollahi, Roxana Zeraati, Emmanouil Giannakakis, Tim Jakob Schäfer et al.ICLR 2024 · 8 citations
- Understanding Adaptive, Multiscale Temporal Integration In Deep Speech Recognition SystemsMenoua Keshishian, Samuel Norman-Haignere, Nima MesgaraniNeurIPS 2021 · 16 citations
