Surrogate Module Learning: Reduce the Gradient Error Accumulation in Training Spiking Neural Networks
Shikuang Deng, Hao Lin, Yuhang Li, Shi Gu
摘要
Spiking neural networks (SNNs) provide an alternative solution to conventional artificial neural networks with energy-saving and high-efficiency characteristics after hardware implantation. However, due to its non-differentiable activation function and the temporally delayed accumulation in outputs, the direct training of SNNs is extraordinarily tough even adopting a surrogate gradient to mimic the backpropagation. For SNN training, this non-differentiability causes the intrinsic gradient error that would be magnified through layerwise backpropagation, especially through multiple layers. In this paper, we propose a novel approach to reducing gradient error from a new perspective called surrogate module learning (SML). Surrogate module learning tries to construct a shortcut path to back-propagate a more accurate gradient to a certain SNN part utilizing the surrogate modules. Then, we develop a new loss function for concurrently training the network and enhancing the surrogate modules' surrogate capacity. We demonstrate that when the outputs of surrogate modules are close to the SNN output, the fraction of the gradient error drops significantly. Our method consistently and significantly enhances the performance of SNNs on all experiment datasets, including CIFAR-10/100, ImageNet, and ES-ImageNet. For example, for spiking ResNet-34 architecture on ImageNet, we increased the SNN accuracy by 3.46%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- CLIF: Complementary Leaky Integrate-and-Fire Neuron for Spiking Neural NetworksYulong Huang, Xiaopeng Lin, Hongwei Ren, Haotian Fu 等ICML 2024 · 被引用 43 次
- TAB: Temporal Accumulated Batch Normalization in Spiking Neural NetworksHaiyan Jiang, Vincent Zoonekynd, Giulia De Masi, Bin Gu 等ICLR 2024 · 被引用 27 次
- Advancing Training Efficiency of Deep Spiking Neural Networks through Rate-based BackpropagationChengting Yu, Lei Liu, Gaoang Wang, Erping Li 等NeurIPS 2024 · 被引用 14 次
- NDOT: Neuronal Dynamics-based Online Training for Spiking Neural NetworksHaiyan Jiang, Giulia De Masi, Huan Xiong, Bin GuICML 2024 · 被引用 13 次
- Autaptic Synaptic Circuit Enhances Spatio-temporal Predictive Learning of Spiking Neural NetworksLihao Wang, Zhaofei YuICML 2024 · 被引用 11 次
它引用的顶会 Paper14
- Be Your Own Teacher: Improve the Performance of Convolutional Neural Networks via Self DistillationLinfeng Zhang, Jiebo Song, Anni Gao, Jingwei Chen 等ICCV 2019 · 被引用 1,069 次
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang 等NeurIPS 2021 · 被引用 857 次
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu 等AAAI 2021 · 被引用 694 次
- Temporal Efficient Training of Spiking Neural Network via Gradient Re-weightingShikuang Deng, Yuhang Li, Shanghang Zhang, Shi GuICLR 2022 · 被引用 361 次
- Differentiable Spike: Rethinking Gradient-Descent for Training Spiking Neural NetworksYuhang Li, Yufei Guo, Shanghang Zhang, Shikuang Deng 等NeurIPS 2021 · 被引用 288 次
相关 Paper
- Towards Memory- and Time-Efficient Backpropagation for Training Spiking Neural NetworksQingyan Meng, Mingqing Xiao, Shen Yan, Yisen Wang 等ICCV 2023 · 被引用 84 次
- SSF: Accelerating Training of Spiking Neural Networks with Stabilized Spiking FlowJingtao Wang, Zengjie Song, Yuxi Wang, Jun Xiao 等ICCV 2023 · 被引用 8 次
- Training Feedback Spiking Neural Networks by Implicit Differentiation on the Equilibrium StateMingqing Xiao, Qingyan Meng, Zongpeng Zhang, Yisen Wang 等NeurIPS 2021 · 被引用 83 次
- Take A Shortcut Back: Mitigating the Gradient Vanishing for Training Spiking Neural NetworksYufei Guo, Yuanpei Chen, Zecheng Hao, Weihang Peng 等NeurIPS 2024 · 被引用 23 次
- Enabling Deep Spiking Neural Networks with Hybrid Conversion and Spike Timing Dependent BackpropagationNitin Rathi, Gopalakrishnan Srinivasan, Priyadarshini Panda, Kaushik RoyICLR 2020 · 被引用 347 次
