Towards High-performance Spiking Transformers from ANN to SNN Conversion
Zihan Huang, Xinyu Shi, Zecheng Hao, Tong Bu, Jianhao Ding, Zhaofei Yu, Tiejun Huang
摘要
Spiking neural networks (SNNs) show great potential due to their energy efficiency, fast processing capabilities, and robustness. There are two main approaches to constructing SNNs. Direct training methods require much memory, while conversion methods offer a simpler and more efficient option. However, current conversion methods mainly focus on converting convolutional neural networks (CNNs) to SNNs. Converting Transformers to SNN is challenging because of the presence of non-linear modules. In this paper, we propose an Expectation Compensation Module to preserve the accuracy of the conversion. The core idea is to use information from the previous T time-steps to calculate the expected output at time-step T. We also propose a Multi-Threshold Neuron and the corresponding Parallel Parameter normalization to address the challenge of large time steps needed for high accuracy, aiming to reduce network latency and power consumption. Our experimental results demonstrate that our approach achieves state-of-the-art performance. For example, we achieve a top-1 accuracy of 88.60% with only a 1% loss in accuracy using 4 time steps while consuming only 35% of the original power of the Transformer. To our knowledge, this is the first successful Artificial Neural Network (ANN) to SNN conversion for Spiking Transformers that achieves high accuracy, low latency, and low power consumption on complex datasets. The source codes of the proposed method are available at https://github.com/h-z-h-cell/Transformer-to-SNN-ECMT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Error Amplification Limits ANN-to-SNN Conversion in Continuous ControlZijie Xu, Zihan Huang, Yiting Dong, Kang Chen 等ICML 2026 · 被引用 2 次
- Training-Free ANN-to-SNN Conversion for High-Performance Spiking TransformersJingya Wang, Xin Deng, Wenjie Wei, Dehao Zhang 等AAAI 2026 · 被引用 1 次
- SpikePack: Enhanced Information Flow in Spiking Neural Networks with High Hardware CompatibilityGuobin Shen, Jindong Li, Tenglong Li, Dongcheng Zhao 等ICCV 2025 · 被引用 1 次
- SpikeVLA: Vision-Language-Action Models with Spiking Neural NetworksRuiqi Song, Dujun Nie, Siyu Teng, Baiyong Ding 等ICML 2026 · 被引用 1 次
- Neural Dynamics Self-Attention for Spiking TransformersDehao Zhang, Fukai Guo, Shuai Wang, Jingya Wang 等ICLR 2026 · 被引用 1 次
它引用的顶会 Paper24
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Learned Step Size quantizationSteven K. Esser, Jeffrey L. McKinstry, Deepika Bablani, Rathinakumar Appuswamy 等ICLR 2020 · 被引用 1,037 次
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang 等NeurIPS 2021 · 被引用 857 次
- Spike-driven TransformerMan Yao, Jiakui Hu, Zhaokun Zhou, Li Yuan 等NeurIPS 2023 · 被引用 368 次
- Optimal ANN-SNN Conversion for High-accuracy and Ultra-low-latency Spiking Neural NetworksTong Bu, Wei Fang, Jianhao Ding, Penglin Dai 等ICLR 2022 · 被引用 272 次
相关 Paper
- SpikeZIP-TF: Conversion is All You Need for Transformer-based SNNKang You, Zekai Xu, Chen Nie, Zhijie Deng 等ICML 2024 · 被引用 20 次
- Generalized Threshold Optimization with Harmony Multi-Threshold Neurons for Accurate ANN-to-SNN ConversionWenhan Zhang, Zihan Huang, Tong Bu, Tiejun Huang 等AAAI 2026
- Efficient ANN-SNN Conversion with Error Compensation LearningChang Liu, Jiangrong Shen, Xuming Ran, Mingkun Xu 等ICML 2025
- Optimized Potential Initialization for Low-Latency Spiking Neural NetworksTong Bu, Jianhao Ding, Zhaofei Yu, Tiejun HuangAAAI 2022 · 被引用 112 次
- TTFSFormer: A TTFS-based Lossless Conversion of Spiking TransformerLusen Zhao, Zihan Huang, Jianhao Ding, Zhaofei YuICML 2025
