Parallel Time Batching: Systolic-Array Acceleration of Sparse Spiking Neural Computation
Jeong-Jun Lee, Wenrui Zhang, Peng Li
摘要
Spiking Neural Networks (SNNs) are brain- inspired computing models incorporating unique temporal dynamics and event-driven processing. Rich dynamics in both space and time offer great challenges and opportunities for efficient processing of sparse spatiotemporal data compared with conventional artificial neural networks (ANNs). Specifically, the additional overheads for handling the added temporal dimension limit the computational capabilities of neuromorphic accelerators. Iterative processing at every time-point with sparse inputs in a temporally sequential manner not only degrades the utilization of the systolic array but also intensifies data movement.In this work, we propose a novel technique and architecture that significantly improve utilization and data movement while efficiently handling temporal sparsity of SNNs on systolic arrays. Unlike time-sequential processing in conventional SNN accelerators, we pack multiple time points into a single time window (TW) and process the computations induced by active synaptic inputs falling under several TWs in parallel, leading to the proposed parallel time batching. It allows weight reuse across multiple time points and enhances the utilization of the systolic array with reduced idling of processing elements, overcoming the irregularity of sparse firing activities. We optimize the granularity of time-domain processing, i.e., the TW size, which significantly impacts the data reuse and utilization. We further boost the utilization efficiency by simultaneously scheduling non-overlapping sparse spiking activities onto the array. The proposed architectures offer a unifying solution for general spiking neural networks with commonly exhibited temporal sparsity, a key challenge in hardware acceleration, delivering 248X energy-delay product (EDP) improvement on average compared to an SNN baseline for accelerating various networks. Compared to ANN based accelerators, our approach improves EDP by 47X on the CIFAR10 dataset.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- LoAS: Fully Temporal-Parallel Dataflow for Dual-Sparse Spiking Neural NetworksRuokai Yin, Youngeun Kim, Di Wu, Priyadarshini PandaMICRO 2024 · 被引用 19 次
- Prosperity: Accelerating Spiking Neural Networks via Product SparsityChiyue Wei, Cong Guo, Feng Cheng, Shiyu Li 等HPCA 2025 · 被引用 14 次
- Phi: Leveraging Pattern-based Hierarchical Sparsity for High-Efficiency Spiking Neural NetworksChiyue Wei, Bowen Duan, Cong Guo, Jingyang Zhang 等ISCA 2025 · 被引用 9 次
- Bishop: Sparsified Bundling Spiking Transformers on Heterogeneous Cores with Error-constrained PruningBoxun Xu, Yuxuan Yin, Vikram Iyer, Peng LiISCA 2025 · 被引用 4 次
- ELSA: An Elastic Snn Inference Architecture for Efficient Neuromorphic ComputingKang You, Chen Nie, Lee Jun Yan, Ziling Wei 等ISCA 2026
相关 Paper
- Single Spike Artificial Neural NetworksRhys Gretsch, Michael Beyeler, Jeremy Lau, Timothy SherwoodISCA 2025
- Temporal Effective Batch Normalization in Spiking Neural NetworksChaoteng Duan, Jianhao Ding, Shiyan Chen, Zhaofei Yu 等NeurIPS 2022 · 被引用 141 次
- Stellar: Energy-Efficient and Low-Latency SNN Algorithm and Hardware Co-Design with Spatiotemporal ComputationRuixin Mao, Lin Tang, Xingyu Yuan, Ye Liu 等HPCA 2024 · 被引用 21 次
- Input-Aware Dynamic Timestep Spiking Neural Networks for Efficient In-Memory ComputingYuhang Li, Abhishek Moitra, Tamar Geller, Priyadarshini PandaDAC 2023 · 被引用 17 次
- Temporal Spike Sequence Learning via Backpropagation for Deep Spiking Neural NetworksWenrui Zhang, Peng LiNeurIPS 2020 · 被引用 264 次
