Masked Spiking Transformer
Ziqing Wang, Yuetong Fang, Jiahang Cao, Qiang Zhang, Zhongrui Wang, Renjing Xu
Abstract
The combination of Spiking Neural Networks (SNNs) and Transformers has attracted significant attention due to their potential for high energy efficiency and high-performance nature. However, existing works on this topic typically rely on direct training, which can lead to suboptimal performance. To address this issue, we propose to leverage the benefits of the ANN-to-SNN conversion method to combine SNNs and Transformers, resulting in significantly improved performance over existing state-of-the-art SNN models. Furthermore, inspired by the quantal synaptic failures observed in the nervous system, which reduce the number of spikes transmitted across synapses, we introduce a novel Masked Spiking Transformer (MST) framework. This incorporates a Random Spike Masking (RSM) method to prune redundant spikes and reduce energy consumption without sacrificing performance. Our experimental results demonstrate that the proposed MST model achieves a significant reduction of 26.8% in power consumption when the masking ratio is 75% while maintaining the same level of performance as the unmasked model. The code is available at: https://github.com/bic-L/Masked-Spiking-Transformer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1b256d7c-884d-488b-b72f-43742ba73e78Cited by top-tier papers32
- QKFormer: Hierarchical Spiking Transformer using Q-K AttentionChenlin Zhou, Han Zhang, Zhaokun Zhou, Liutao Yu et al.NeurIPS 2024 · 126 citations
- Autonomous Driving with Spiking Neural NetworksRuijie Zhu, Ziqing Wang, Leilani Gilpin, Jason EshraghianNeurIPS 2024 · 35 citations
- Spiking Meets Attention: Efficient Remote Sensing Image Super-Resolution with Attention Spiking Neural NetworksYi Xiao, Qiangqiang Yuan, Kui Jiang, Wenke Huang et al.NeurIPS 2025 · 25 citations
- High-Performance Temporal Reversible Spiking Neural Networks with O(L) Training Memory and O(1) Inference CostJiakui Hu, Man Yao, Xuerui Qiu, Yuhong Chou et al.ICML 2024 · 24 citations
- SpikedAttention: Training-Free and Fully Spike-Driven Transformer-to-SNN Conversion with Winner-Oriented Spike Shift for Softmax OperationSangwoo Hwang, Seunghyun Lee, Dahoon Park, Donghun Lee et al.NeurIPS 2024 · 23 citations
Builds on15
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang et al.NeurIPS 2021 · 857 citations
- Incorporating Learnable Membrane Time Constant to Enhance Learning of Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Timothée Masquelier et al.ICCV 2021 · 731 citations
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu et al.AAAI 2021 · 694 citations
Related papers
- Towards High-performance Spiking Transformers from ANN to SNN ConversionZihan Huang, Xinyu Shi, Zecheng Hao, Tong Bu et al.ACM MM 2024 · 17 citations
- Spiking Transformer: Introducing Accurate Addition-Only Spiking Self-Attention for TransformerYufei Guo, Xiaode Liu, Yuanpei Chen, Weihang Peng et al.CVPR 2025
- SpikeZIP-TF: Conversion is All You Need for Transformer-based SNNKang You, Zekai Xu, Chen Nie, Zhijie Deng et al.ICML 2024 · 20 citations
- TTFSFormer: A TTFS-based Lossless Conversion of Spiking TransformerLusen Zhao, Zihan Huang, Jianhao Ding, Zhaofei YuICML 2025
- LM-HT SNN: Enhancing the Performance of SNN to ANN Counterpart through Learnable Multi-hierarchical Threshold ModelZecheng Hao, Xinyu Shi, Yujia Liu, Zhaofei Yu et al.NeurIPS 2024 · 16 citations
