Training Deep Spiking Neural Networks without Normalization
Xinyu Shi, Zhaofei Yu
Abstract
The training of deep Spiking Neural Networks (SNNs) has traditionally relied on Batch Normalization (BN), which stabilizes input currents and gradients during training. However, BN is not a universal solution. It is unsuitable for variablelength tasks and scenarios with reduced batch size, constraining the development of deep SNNs, where removing BN typically causes the training to fail to converge. This dependence stems not from a fundamental necessity of BN but from the current lack of reasonable initialization methods for SNNs. This paper addresses this core limitation by proposing SpikeInit, a novel initialization framework for SNNs. By modeling the response curve and gradient of spiking layers, SpikeInit initializes the weights and shape parameters of surrogate gradients to maintain stable firing rates during forward propagation and stable gradient magnitudes during backpropagation. Extensive experiments demonstrate that deep SNNs with SpikeInit can be trained stably without normalization and achieve superior performance compared to their normalized counterparts under identical settings. Furthermore, we demonstrate the scalability of SpikeInit by successfully training an ultra-deep, 1000-layer SNN without normalization. Our work provides a foundational step toward large-scale normalizationfree SNN, liberating SNN design from the constraints of normalization. Codes are available at: https://github.com/xyshi2000/SpikeInit
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on15
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang et al.NeurIPS 2021 · 857 citations
- Incorporating Learnable Membrane Time Constant to Enhance Learning of Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Timothée Masquelier et al.ICCV 2021 · 731 citations
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu et al.AAAI 2021 · 694 citations
- High-Performance Large-Scale Image Recognition Without NormalizationAndy Brock, Soham De, Samuel L. Smith, Karen SimonyanICML 2021 · 613 citations
- Spike-driven TransformerMan Yao, Jiakui Hu, Zhaokun Zhou, Li Yuan et al.NeurIPS 2023 · 368 citations
Related papers
- Training Deep Normalization-Free Spiking Neural Networks with Lateral InhibitionPeiyu Liu, Jianhao Ding, Zhaofei YuICLR 2026 · 1 citation
- Temporal Effective Batch Normalization in Spiking Neural NetworksChaoteng Duan, Jianhao Ding, Shiyan Chen, Zhaofei Yu et al.NeurIPS 2022 · 141 citations
- Batch Normalization Biases Residual Blocks Towards the Identity Function in Deep NetworksSoham De, Samuel L. SmithNeurIPS 2020 · 173 citations
- UniSparse: Combining Weight Pruning and Spike Sparsification in Spiking Neural NetworksXinyu Shi, Tong Bu, Zhaofei YuICML 2026
- Membrane Potential Batch Normalization for Spiking Neural NetworksYufei Guo, Yuhan Zhang, Yuanpei Chen, Weihang Peng et al.ICCV 2023 · 62 citations
