TPipe: Efficient Spiking Transformer Training with Time Parallelism and Asynchronous Pipeline
Yubing Bao, ZhiHui Lu, Qiang Duan, Changze Lv, Xin Du, Zeyi Deng, Jingqi Feng, Sen Liu, Yang Chen, Xin Wang
摘要
Spiking Transformers, which couple Spiking Neural Networks (SNNs) with Transformer architectures, promise low-energy inference and strong accuracy. However, their training is hindered by the intrinsic temporal dynamics of SNNs, which cause severe memory and synchronization bottlenecks. Through an in-depth analysis of the SpikeFormer architecture and training workload, we uncover local dependency in its training process, which enables time parallelism. We present TPipe, the first unified framework for efficient SpikeFormer training that integrates time, data, and pipeline parallelism. We formalize the Automated SpikeFormer Parallelism (ASP) problem to jointly optimize parallelism and placement strategies. We further formalize the SpikeFormer Scheduling (SS) problem, which eliminates pipeline bubbles via fine-grained, sub–micro-batch asynchronous scheduling. We prove that TPipe achieves bounded speed-up over synchronous baselines while preserving training correctness. Experiments on multiple SpikeFormer models and datasets show that TPipe attains 1.1–3.8× speed-up over default pipeline baselines while maintaining accuracy.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- AdaptPipe: Mitigating Runtime Bubbles via Granularity-Adaptive Scheduling under Memory ConstraintsYumeng Cui, Jessie Hui Wang, Najila Liu, Ling Deng 等INFOCOM 2026
- HelixPipe: Efficient Distributed Training of Long Sequence Transformers with Attention Parallel Pipeline ParallelismGeng Zhang, Shenggan Cheng, Xuanlei Zhao, Ziming Liu 等PPoPP 2026 · 被引用 3 次
- TEFormer: Structured Bidirectional Temporal Enhancement Modeling in Spiking TransformersSicheng Shen, Mingyang Lv, Bing Han, Dongcheng Zhao 等ICML 2026 · 被引用 1 次
- FlexPipe: Maximizing Training Efficiency for Transformer-based Models with Variable-Length InputsHairui Zhao, Qi Tian, Hongliang Li, Zizhong ChenUSENIX ATC 2025 · 被引用 6 次
- Towards Efficient Spiking Transformer: a Token Sparsification Framework for Training and Inference AccelerationZhengyang Zhuge, Peisong Wang, Xingting Yao, Jian ChengICML 2024 · 被引用 6 次
