Differentiable Spike: Rethinking Gradient-Descent for Training Spiking Neural Networks
Yuhang Li, Yufei Guo, Shanghang Zhang, Shikuang Deng, Yongqing Hai, Shi Gu
摘要
Spiking neural networks (SNNs) have emerged as a biology-inspired method mimicking the spiking nature of brain neurons. This biomimicry derives SNNs' energy efficiency of inference on neuromorphic hardware. However, it also causes an intrinsic disadvantage in training high-performing SNNs from scratch since the discrete spike prohibits the gradient calculation. To overcome this issue, the surrogate gradient (SG) approach has been proposed as a continuous relaxation. Yet the heuristic choice of SG leaves it vacant how the SG benefits the SNN training. In this work, we first theoretically study the gradient descent problem in SNN training and introduce finite difference gradient to quantitatively analyze the training behavior of SNN. Based on the introduced finite difference gradient, we propose a new family of Differentiable Spike (Dspike) functions that can adaptively evolve during training to find the optimal shape and smoothness for gradient estimation. Extensive experiments over several popular network structures show that training SNN with Dspike consistently outperforms the state-of-the-art training methods. For example, on the CIFAR10-DVS classification task, we can train a spiking ResNet-18 and achieve 75.4% top-1 accuracy with 10 time steps.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper94
- Spike-driven TransformerMan Yao, Jiakui Hu, Zhaokun Zhou, Li Yuan 等NeurIPS 2023 · 被引用 368 次
- Temporal Efficient Training of Spiking Neural Network via Gradient Re-weightingShikuang Deng, Yuhang Li, Shanghang Zhang, Shi GuICLR 2022 · 被引用 361 次
- GLIF: A Unified Gated Leaky Integrate-and-Fire Neuron for Spiking Neural NetworksXingting Yao, Fanrong Li, Zitao Mo, Jian ChengNeurIPS 2022 · 被引用 175 次
- Temporal Effective Batch Normalization in Spiking Neural NetworksChaoteng Duan, Jianhao Ding, Shiyan Chen, Zhaofei Yu 等NeurIPS 2022 · 被引用 141 次
- IM-Loss: Information Maximization Loss for Spiking Neural NetworksYufei Guo, Yuanpei Chen, Liwen Zhang, Xiaode Liu 等NeurIPS 2022 · 被引用 129 次
它引用的顶会 Paper9
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu 等AAAI 2021 · 被引用 694 次
- BRECQ: Pushing the Limit of Post-Training Quantization by Block ReconstructionYuhang Li, Ruihao Gong, Xu Tan, Yang Yang 等ICLR 2021 · 被引用 619 次
- Enabling Deep Spiking Neural Networks with Hybrid Conversion and Spike Timing Dependent BackpropagationNitin Rathi, Gopalakrishnan Srinivasan, Priyadarshini Panda, Kaushik RoyICLR 2020 · 被引用 347 次
- Additive Powers-of-Two Quantization: An Efficient Non-uniform Discretization for Neural NetworksYuhang Li, Xin Dong, Wei WangICLR 2020 · 被引用 315 次
- Temporal Spike Sequence Learning via Backpropagation for Deep Spiking Neural NetworksWenrui Zhang, Peng LiNeurIPS 2020 · 被引用 264 次
相关 Paper
- Surrogate Module Learning: Reduce the Gradient Error Accumulation in Training Spiking Neural NetworksShikuang Deng, Hao Lin, Yuhang Li, Shi GuICML 2023 · 被引用 36 次
- Training Feedback Spiking Neural Networks by Implicit Differentiation on the Equilibrium StateMingqing Xiao, Qingyan Meng, Zongpeng Zhang, Yisen Wang 等NeurIPS 2021 · 被引用 83 次
- Training High-Performance Low-Latency Spiking Neural Networks by Differentiation on Spike RepresentationQingyan Meng, Mingqing Xiao, Shen Yan, Yisen Wang 等CVPR 2022 · 被引用 114 次
- DeepTAGE: Deep Temporal-Aligned Gradient Enhancement for Optimizing Spiking Neural NetworksWei Liu, Li Yang, Mingxuan Zhao, Shuxun Wang 等ICLR 2025
- Take A Shortcut Back: Mitigating the Gradient Vanishing for Training Spiking Neural NetworksYufei Guo, Yuanpei Chen, Zecheng Hao, Weihang Peng 等NeurIPS 2024 · 被引用 23 次
