SMixer: Rethinking Efficient-Training and Event-Driven SNNs
Yijie Lu, Xinhao Luo, Yixing Zhang, Zhiyan Wang, Wentao Li, YanhanWang, Zhi Liu, Zhaokun Zhou, Guoqi Li
摘要
Spiking Neural Networks (SNNs) offer a promising, energy-efficient paradigm for computation, but their practical application is hindered by challenges in architecture design and training costs. For example, Spiking ResNet exhibits relatively low performance, whereas high-performance Spiking Transformers are not truly event driven and cannot be implemented on asynchronous chips. Moreover, the intrinsic time steps and neuron state dynamics result in a substantial computational overhead for training SNNs on GPUs. In response to these problems, we discuss rational architectural design for SNNs and argue that such designs should exhibit three key characteristics: operations fully supported by asynchronous scenarios, low training overhead and competitive performance. In light of this, we adopt the event-driven friendly Spiking Mixer (SMixer) as the foundational architecture and develop a spike feature Spatial-Temporal Pruning (STP) framework with a high pruning ratio and no trainable parameters to reduce the training overhead. Based on a statistical analysis of sparse spike features, STP eliminates redundant spike features across both spatial and temporal dimensions, thereby reducing the input features and computational load during training. It adaptively selects the most salient spike tokens spatially and dynamically constrains neuron firing rates temporally. By leveraging STP and architectural adaptation, SMixer accelerates training while ensuring a fully event-driven characteristics and maintaining competitive performance, offering valuable insights for the design of efficient, event-driven SNNs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Spike-driven TransformerMan Yao, Jiakui Hu, Zhaokun Zhou, Li Yuan 等NeurIPS 2023 · 被引用 368 次
- Drawing Early-Bird Tickets: Toward More Efficient Training of Deep NetworksHaoran You, Chaojian Li, Pengfei Xu, Yonggan Fu 等ICLR 2020 · 被引用 282 次
- Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic ChipsMan Yao, Jiakui Hu, Tianxiang Hu, Yifan Xu 等ICLR 2024 · 被引用 154 次
- QKFormer: Hierarchical Spiking Transformer using Q-K AttentionChenlin Zhou, Han Zhang, Zhaokun Zhou, Liutao Yu 等NeurIPS 2024 · 被引用 126 次
- Spikformer: When Spiking Neural Network Meets TransformerZhaokun Zhou, Yuesheng Zhu, Chao He, Yaowei Wang 等ICLR 2023 · 被引用 103 次
相关 Paper
- Spiking Token Mixer: An event-driven friendly Former structure for spiking neural networksShikuang Deng, Yuhang Wu, Kangrui Du, Shi GuNeurIPS 2024 · 被引用 9 次
- TP-Spikformer: Token Pruned Spiking TransformerWenjie Wei, Xiaolong Zhou, Malu Zhang, Ammar Belatreche 等ICLR 2026 · 被引用 6 次
- Spikingformer: A Key Foundation Model for Spiking Neural NetworksChenlin Zhou, Liutao Yu, Zhaokun Zhou, Han Zhang 等AAAI 2026 · 被引用 4 次
- Temporal Flexibility in Spiking Neural Networks: Towards Generalization Across Time Steps and Deployment FriendlinessKangrui Du, Yuhang Wu, Shikuang Deng, Shi GuICLR 2025
- Temporal Interaction in Spiking Transformers with Multi-Delay MixerKexin Shi, Hanwen Liu, Zeyang Song, Yang Liu 等CVPR 2026
