Positional Encoding for Spiking Transformers
Zijian Zhou, Yu Liang, Honglin Cao, Ammar Belatreche, Jieyuan Zhang, Wenjie Wei, Shuai Wang, Malu Zhang, Yang Yang, Haizhou Li
摘要
Transformer-based Spiking Neural Networks (SNNs) have recently emerged as a promising paradigm to sequential modeling, combining the strong representational capabilities of Transformers with the sparse spike-driven computation of SNNs. Within such position-agnostic architectures, positional encoding is critical for injecting order information, allowing the model to distinguish token positions, capture sequential dependencies, and represent relative relationships among tokens. However, existing positional encoding methods for SNNs are largely inherited from ANNs and, in doing so, undermine the spike-driven computational properties that are central to spiking transformers. To address this limitation, we propose the Spiking Positional Encoding (SPE), a method designed specifically for Spiking Transformers, aimed at encoding relative positional information while preserving both spike-driven computation and the linear complexity of spiking self-attention. The core component of SPE is the Positional Encoding Leaky Integrate-and-Fire (PE-LIF) neuron, which incorporates position-dependent signals into neuronal thresholds and implicitly propagates this information through spike trains via continuous firing and membrane potential reset dynamics. Extensive experiments on thirteen NLP benchmarks demonstrate that SPE consistently outperforms existing SNN positional encoding methods, strengthens the sequence modeling capability, and improves energy efficiency without introducing additional trainable parameters. Code is available at https://github.com/CayleyZ/SPE.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 被引用 2,665 次
- Spike-driven TransformerMan Yao, Jiakui Hu, Zhaokun Zhou, Li Yuan 等NeurIPS 2023 · 被引用 368 次
- Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic ChipsMan Yao, Jiakui Hu, Tianxiang Hu, Yifan Xu 等ICLR 2024 · 被引用 154 次
- QKFormer: Hierarchical Spiking Transformer using Q-K AttentionChenlin Zhou, Han Zhang, Zhaokun Zhou, Liutao Yu 等NeurIPS 2024 · 被引用 126 次
相关 Paper
- Toward Relative Positional Encoding in Spiking TransformersChangze Lv, Yansen Wang, Dongqi Han, Yifei Shen 等NeurIPS 2025 · 被引用 8 次
- SpikF: Spiking Fourier Network for Efficient Long-term PredictionWenjie Wu, Dexuan Huo, Hong ChenICML 2025
- Complex Dynamic Neurons Improved Spiking Transformer Network for Efficient Automatic Speech RecognitionQingyu Wang, Tielin Zhang, Minglun Han, Yi Wang 等AAAI 2023 · 被引用 40 次
- Advancing Spiking Neural Networks for Sequential Modeling with Central Pattern GeneratorsChangze Lv, Dongqi Han, Yansen Wang, Xiaoqing Zheng 等NeurIPS 2024 · 被引用 10 次
- AdaS: Adaptive Gradient Descent for Spiking TransformersZijian Zhou, Honglin Cao, Ammar Belatreche, Wenjie Wei 等ICML 2026
