Deep Reinforcement Learning with Time-Scale Invariant Memory
Md Rysul Kabir, James Mochizuki-Freeman, Zoran Tiganj
Abstract
The ability to estimate temporal relationships is critical for both animals and artificial agents. Cognitive science and neuroscience provide remarkable insights into behavioral and neural aspects of temporal credit assignment. In particular, scale invariance of learning dynamics, observed in behavior and supported by neural data, is one of the key principles that governs animal perception: proportional rescaling of temporal relationships does not alter the overall learning efficiency. Here we integrate a computational neuroscience model of scale invariant memory into deep reinforcement learning (RL) agents. We first provide a theoretical analysis and then demonstrate through experiments that such agents can learn robustly across a wide range of temporal scales, unlike agents built with commonly used recurrent memory architectures such as LSTM. This result illustrates that incorporating computational principles from neuroscience and cognitive science into deep neural networks can enhance adaptability to complex temporal dynamics, mirroring some of the core properties of human learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a7f131c9-ca13-4072-80f1-e1b074dc5e47Builds on4
- Relating transformers to models and neural representations of the hippocampal formationJames C. R. Whittington, Joseph Warren, Tim E. J. BehrensICLR 2022 · 110 citations
- No Free Lunch from Deep Learning in Neuroscience: A Case Study through Models of the Entorhinal-Hippocampal CircuitRylan Schaeffer, Mikail Khona, Ila FieteNeurIPS 2022 · 81 citations
- A deep convolutional neural network that is invariant to time rescalingBrandon G. Jacques, Zoran Tiganj, Aakash Sarkar, Marc W. Howard et al.ICML 2022 · 10 citations
- DeepSITH: Efficient Learning via Decomposition of What and When Across Time ScalesBrandon G. Jacques, Zoran Tiganj, Marc W. Howard, Per B. SederbergNeurIPS 2021 · 9 citations
Related papers
- The least-control principle for local learning at equilibriumAlexander Meulemans, Nicolas Zucchet, Seijin Kobayashi, Johannes von Oswald et al.NeurIPS 2022 · 32 citations
- A Recurrent Neural Circuit Mechanism of Temporal-scaling Equivariant RepresentationJunfeng Zuo, Xiao Liu, Ying Nian Wu, Si Wu et al.NeurIPS 2023 · 6 citations
- Flexible inference for animal learning rules using neural networksYuhan Helena Liu, Victor Geadah, Jonathan W. PillowNeurIPS 2025
- Deep RL Needs Deep Behavior Analysis: Exploring Implicit Planning by Model-Free Agents in Open-Ended EnvironmentsRiley Simmons-Edler, Ryan Paul Badman, Felix Baastad Berg, Raymond Chua et al.NeurIPS 2025 · 6 citations
- Inverse Rational Control with Partially Observable Continuous Nonlinear DynamicsMinhae Kwon, Saurabh Daptardar, Paul R. Schrater, Xaq PitkowNeurIPS 2020 · 46 citations
