Deep Reinforcement Learning with Time-Scale Invariant Memory
Md Rysul Kabir, James Mochizuki-Freeman, Zoran Tiganj
摘要
The ability to estimate temporal relationships is critical for both animals and artificial agents. Cognitive science and neuroscience provide remarkable insights into behavioral and neural aspects of temporal credit assignment. In particular, scale invariance of learning dynamics, observed in behavior and supported by neural data, is one of the key principles that governs animal perception: proportional rescaling of temporal relationships does not alter the overall learning efficiency. Here we integrate a computational neuroscience model of scale invariant memory into deep reinforcement learning (RL) agents. We first provide a theoretical analysis and then demonstrate through experiments that such agents can learn robustly across a wide range of temporal scales, unlike agents built with commonly used recurrent memory architectures such as LSTM. This result illustrates that incorporating computational principles from neuroscience and cognitive science into deep neural networks can enhance adaptability to complex temporal dynamics, mirroring some of the core properties of human learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Relating transformers to models and neural representations of the hippocampal formationJames C. R. Whittington, Joseph Warren, Tim E. J. BehrensICLR 2022 · 被引用 110 次
- No Free Lunch from Deep Learning in Neuroscience: A Case Study through Models of the Entorhinal-Hippocampal CircuitRylan Schaeffer, Mikail Khona, Ila FieteNeurIPS 2022 · 被引用 81 次
- A deep convolutional neural network that is invariant to time rescalingBrandon G. Jacques, Zoran Tiganj, Aakash Sarkar, Marc W. Howard 等ICML 2022 · 被引用 10 次
- DeepSITH: Efficient Learning via Decomposition of What and When Across Time ScalesBrandon G. Jacques, Zoran Tiganj, Marc W. Howard, Per B. SederbergNeurIPS 2021 · 被引用 9 次
相关 Paper
- The least-control principle for local learning at equilibriumAlexander Meulemans, Nicolas Zucchet, Seijin Kobayashi, Johannes von Oswald 等NeurIPS 2022 · 被引用 32 次
- A Recurrent Neural Circuit Mechanism of Temporal-scaling Equivariant RepresentationJunfeng Zuo, Xiao Liu, Ying Nian Wu, Si Wu 等NeurIPS 2023 · 被引用 6 次
- Flexible inference for animal learning rules using neural networksYuhan Helena Liu, Victor Geadah, Jonathan W. PillowNeurIPS 2025
- Deep RL Needs Deep Behavior Analysis: Exploring Implicit Planning by Model-Free Agents in Open-Ended EnvironmentsRiley Simmons-Edler, Ryan Paul Badman, Felix Baastad Berg, Raymond Chua 等NeurIPS 2025 · 被引用 6 次
- Inverse Rational Control with Partially Observable Continuous Nonlinear DynamicsMinhae Kwon, Saurabh Daptardar, Paul R. Schrater, Xaq PitkowNeurIPS 2020 · 被引用 46 次
