GRSN: Gated Recurrent Spiking Neurons for POMDPs and MARL
Lang Qin, Ziming Wang, Runhao Jiang, Rui Yan, Huajin Tang
Abstract
Spiking neural networks (SNNs) are widely applied in various fields due to their energy-efficient and fast-inference capabilities. Applying SNNs to reinforcement learning (RL) can significantly reduce the computational resource requirements for agents and improve the algorithm's performance under resource-constrained conditions. However, in current spiking reinforcement learning (SRL) algorithms, the simulation results of multiple time steps can only correspond to a single-step decision in RL. This is quite different from the real temporal dynamics in the brain and also fails to fully exploit the capacity of SNNs to process temporal data. In order to address this temporal mismatch issue and further take advantage of the inherent temporal dynamics of spiking neurons, we propose a novel temporal alignment paradigm (TAP) that leverages the single-step update of spiking neurons to accumulate historical state information in RL and introduces gated units to enhance the memory capacity of spiking neurons. Experimental results show that our method can solve partially observable Markov decision processes (POMDPs) and multi-agent cooperation problems with similar performance as recurrent neural networks (RNNs) but with about 50% power consumption.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on3
- Recurrent Model-Free RL Can Be a Strong Baseline for Many POMDPsTianwei Ni, Benjamin Eysenbach, Ruslan SalakhutdinovICML 2022 · 162 citations
- Variational Recurrent Models for Solving Partially Observable Control TasksDongqi Han, Kenji Doya, Jun TaniICLR 2020 · 75 citations
- Adaptive Smoothing Gradient Learning for Spiking Neural NetworksZiming Wang, Runhao Jiang, Shuang Lian, Rui Yan et al.ICML 2023 · 69 citations
Related papers
- Multi-Sacle Dynamic Coding Improved Spiking Actor Network for Reinforcement LearningDuzhen Zhang, Tielin Zhang, Shuncheng Jia, Bo XuAAAI 2022 · 45 citations
- Resolving the Timestep Scaling Paradox in Spiking Neural Networks with a Timestep-Scalable Neuron ModelBinghao Ye, Wenjuan Li, Dengfeng Xue, Bing Li et al.ICML 2026
- Proxy Target: Bridging the Gap Between Discrete Spiking Neural Networks and Continuous ControlZijie Xu, Tong Bu, Zecheng Hao, Jianhao Ding et al.NeurIPS 2025 · 10 citations
- DeepTAGE: Deep Temporal-Aligned Gradient Enhancement for Optimizing Spiking Neural NetworksWei Liu, Li Yang, Mingxuan Zhao, Shuxun Wang et al.ICLR 2025
- TS-SNN: Temporal Shift Module for Spiking Neural NetworksKairong Yu, Tianqing Zhang, Qi Xu, Gang Pan et al.ICML 2025
