SMixer: Rethinking Efficient-Training and Event-Driven SNNs
Yijie Lu, Xinhao Luo, Yixing Zhang, Zhiyan Wang, Wentao Li, YanhanWang, Zhi Liu, Zhaokun Zhou, Guoqi Li
Abstract
Spiking Neural Networks (SNNs) offer a promising, energy-efficient paradigm for computation, but their practical application is hindered by challenges in architecture design and training costs. For example, Spiking ResNet exhibits relatively low performance, whereas high-performance Spiking Transformers are not truly event driven and cannot be implemented on asynchronous chips. Moreover, the intrinsic time steps and neuron state dynamics result in a substantial computational overhead for training SNNs on GPUs. In response to these problems, we discuss rational architectural design for SNNs and argue that such designs should exhibit three key characteristics: operations fully supported by asynchronous scenarios, low training overhead and competitive performance. In light of this, we adopt the event-driven friendly Spiking Mixer (SMixer) as the foundational architecture and develop a spike feature Spatial-Temporal Pruning (STP) framework with a high pruning ratio and no trainable parameters to reduce the training overhead. Based on a statistical analysis of sparse spike features, STP eliminates redundant spike features across both spatial and temporal dimensions, thereby reducing the input features and computational load during training. It adaptively selects the most salient spike tokens spatially and dynamically constrains neuron firing rates temporally. By leveraging STP and architectural adaptation, SMixer accelerates training while ensuring a fully event-driven characteristics and maintaining competitive performance, offering valuable insights for the design of efficient, event-driven SNNs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50d1bfd0-3b3e-4f2e-b8d7-3883170b07c3Builds on14
- Spike-driven TransformerMan Yao, Jiakui Hu, Zhaokun Zhou, Li Yuan et al.NeurIPS 2023 · 368 citations
- Drawing Early-Bird Tickets: Toward More Efficient Training of Deep NetworksHaoran You, Chaojian Li, Pengfei Xu, Yonggan Fu et al.ICLR 2020 · 282 citations
- Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic ChipsMan Yao, Jiakui Hu, Tianxiang Hu, Yifan Xu et al.ICLR 2024 · 154 citations
- QKFormer: Hierarchical Spiking Transformer using Q-K AttentionChenlin Zhou, Han Zhang, Zhaokun Zhou, Liutao Yu et al.NeurIPS 2024 · 126 citations
- Spikformer: When Spiking Neural Network Meets TransformerZhaokun Zhou, Yuesheng Zhu, Chao He, Yaowei Wang et al.ICLR 2023 · 103 citations
Related papers
- Spiking Token Mixer: An event-driven friendly Former structure for spiking neural networksShikuang Deng, Yuhang Wu, Kangrui Du, Shi GuNeurIPS 2024 · 9 citations
- TP-Spikformer: Token Pruned Spiking TransformerWenjie Wei, Xiaolong Zhou, Malu Zhang, Ammar Belatreche et al.ICLR 2026 · 6 citations
- Spikingformer: A Key Foundation Model for Spiking Neural NetworksChenlin Zhou, Liutao Yu, Zhaokun Zhou, Han Zhang et al.AAAI 2026 · 4 citations
- Temporal Flexibility in Spiking Neural Networks: Towards Generalization Across Time Steps and Deployment FriendlinessKangrui Du, Yuhang Wu, Shikuang Deng, Shi GuICLR 2025
- Temporal Interaction in Spiking Transformers with Multi-Delay MixerKexin Shi, Hanwen Liu, Zeyang Song, Yang Liu et al.CVPR 2026
