Direct Training of SNN using Local Zeroth Order Method
Bhaskar Mukhoty, Velibor Bojkovic, William de Vazelhes, Xiaohan Zhao, Giulia De Masi, Huan Xiong, Bin Gu
摘要
Spiking neural networks are becoming increasingly popular for their low energy requirement in real-world tasks with accuracy comparable to traditional ANNs. SNN training algorithms face the loss of gradient information and non-differentiability due to the Heaviside function in minimizing the model loss over model parameters. To circumvent this problem, the surrogate method employs a differentiable approximation of the Heaviside function in the backward pass, while the forward pass continues to use the Heaviside as the spiking function. We propose to use the zeroth-order technique at the local or neuron level in training SNNs, motivated by its regularizing and potential energy-efficient effects and establish a theoretical connection between it and the existing surrogate methods. We perform experimental validation of the technique on standard static datasets (CIFAR-10, CIFAR-100, ImageNet-100) and neuromorphic datasets (DVS-CIFAR-10, DVS-Gesture, N-Caltech-101, NCARS) and obtain results that offer improvement over the state-of-the-art results. The proposed method also lends itself to efficient implementations of the back-propagation method, which could provide 3-4 times overall speedup in training time. The code is available at https://github.com/BhaskarMukhoty/LocalZO .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- CLIF: Complementary Leaky Integrate-and-Fire Neuron for Spiking Neural NetworksYulong Huang, Xiaopeng Lin, Hongwei Ren, Haotian Fu 等ICML 2024 · 被引用 43 次
- Enhancing Training of Spiking Neural Network with Stochastic LatencySrinivas Anumasa, Bhaskar Mukhoty, Velibor Bojkovic, Giulia De Masi 等AAAI 2024 · 被引用 17 次
- Online Pseudo-Zeroth-Order Training of Neuromorphic Spiking Neural NetworksMingqing Xiao, Qingyan Meng, Zongpeng Zhang, Di He 等ICLR 2026 · 被引用 2 次
- SAFA-SNN: Sparsity-Aware On-Device Few-Shot Class-Incremental Learning with Fast-Adaptive Structure of Spiking Neural NetworkHuijing Zhang, Muyang Cao, Linshan Jiang, Xin Du 等ICLR 2026 · 被引用 1 次
- Towards Reliable Evaluation of Adversarial Robustness for Spiking Neural NetworksJihang Wang, Dongcheng Zhao, Ruolin Chen, Qian Zhang 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper11
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang 等NeurIPS 2021 · 被引用 857 次
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu 等AAAI 2021 · 被引用 694 次
- Spiking-YOLO: Spiking Neural Network for Energy-Efficient Object DetectionSei Joon Kim, Seongsik Park, Byunggook Na, Sungroh YoonAAAI 2020 · 被引用 512 次
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 被引用 427 次
- Temporal Efficient Training of Spiking Neural Network via Gradient Re-weightingShikuang Deng, Yuhang Li, Shanghang Zhang, Shi GuICLR 2022 · 被引用 361 次
相关 Paper
- Differentiable Spike: Rethinking Gradient-Descent for Training Spiking Neural NetworksYuhang Li, Yufei Guo, Shanghang Zhang, Shikuang Deng 等NeurIPS 2021 · 被引用 288 次
- Towards Memory- and Time-Efficient Backpropagation for Training Spiking Neural NetworksQingyan Meng, Mingqing Xiao, Shen Yan, Yisen Wang 等ICCV 2023 · 被引用 84 次
- Training Feedback Spiking Neural Networks by Implicit Differentiation on the Equilibrium StateMingqing Xiao, Qingyan Meng, Zongpeng Zhang, Yisen Wang 等NeurIPS 2021 · 被引用 83 次
- Enabling Deep Spiking Neural Networks with Hybrid Conversion and Spike Timing Dependent BackpropagationNitin Rathi, Gopalakrishnan Srinivasan, Priyadarshini Panda, Kaushik RoyICLR 2020 · 被引用 347 次
- DeepTAGE: Deep Temporal-Aligned Gradient Enhancement for Optimizing Spiking Neural NetworksWei Liu, Li Yang, Mingxuan Zhao, Shuxun Wang 等ICLR 2025
