SpinalFlow: An Architecture and Dataflow Tailored for Spiking Neural Networks
Surya Narayanan, Karl Taht, Rajeev Balasubramonian, Edouard Giacomin, Pierre-Emmanuel Gaillardon
摘要
Spiking neural networks (SNNs) are expected to be part of the future AI portfolio, with heavy investment from industry and government, e.g., IBM TrueNorth, Intel Loihi. While Artificial Neural Network (ANN) architectures have taken large strides, few works have targeted SNN hardware efficiency. Our analysis of SNN baselines shows that at modest spike rates, SNN implementations exhibit significantly lower efficiency than accelerators for ANNs. This is primarily because SNN dataflows must consider neuron potentials for several ticks, introducing a new data structure and a new dimension to the reuse pattern. We introduce a novel SNN architecture, SpinalFlow, that processes a compressed, time-stamped, sorted sequence of input spikes. It adopts an ordering of computations such that the outputs of a network layer are also compressed, time-stamped, and sorted. All relevant computations for a neuron are performed in consecutive steps to eliminate neuron potential storage overheads. Thus, with better data reuse, we advance the energy efficiency of SNN accelerators by an order of magnitude. Even though the temporal aspect in SNNs prevents the exploitation of some reuse patterns that are more easily exploited in ANNs, at 4-bit input resolution and 90% input sparsity, SpinalFlow reduces average energy by 1.8×, compared to a 4-bit Eyeriss baseline. These improvements are seen for a range of networks and sparsity/resolution levels; SpinalFlow consumes 5× less energy and 5.4× less time than an 8-bit version of Eyeriss. We thus show that, depending on the level of observed sparsity, SNN architectures can be competitive with ANN architectures in terms of latency and energy for inference, thus lowering the barrier for practical deployment in scenarios demanding real-time learning.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper10
- Toward Robust Spiking Neural Network Against Adversarial PerturbationLing Liang, Kaidi Xu, Xing Hu, Lei Deng 等NeurIPS 2022 · 被引用 28 次
- SATO: spiking neural network acceleration via temporal-oriented dataflow and architectureFangxin Liu, Wenbo Zhao, Zongwu Wang, Yongbiao Chen 等DAC 2022 · 被引用 20 次
- LoAS: Fully Temporal-Parallel Dataflow for Dual-Sparse Spiking Neural NetworksRuokai Yin, Youngeun Kim, Di Wu, Priyadarshini PandaMICRO 2024 · 被引用 19 次
- A time-to-first-spike coding and conversion aware training for energy-efficient deep spiking neural network processor designDongwoo Lew, Kyungchul Lee, Jongsun ParkDAC 2022 · 被引用 18 次
- Prosperity: Accelerating Spiking Neural Networks via Product SparsityChiyue Wei, Cong Guo, Feng Cheng, Shiyu Li 等HPCA 2025 · 被引用 14 次
相关 Paper
- Stellar: Energy-Efficient and Low-Latency SNN Algorithm and Hardware Co-Design with Spatiotemporal ComputationRuixin Mao, Lin Tang, Xingyu Yuan, Ye Liu 等HPCA 2024 · 被引用 21 次
- SpikePack: Enhanced Information Flow in Spiking Neural Networks with High Hardware CompatibilityGuobin Shen, Jindong Li, Tenglong Li, Dongcheng Zhao 等ICCV 2025 · 被引用 1 次
- ELSA: An Elastic Snn Inference Architecture for Efficient Neuromorphic ComputingKang You, Chen Nie, Lee Jun Yan, Ziling Wei 等ISCA 2026
- Accelerating Linear Recurrent Neural Networks for the Edge with Unstructured SparsityAlessandro Pierro, Steven Abreu, Jonathan Timcheck, Philipp Stratmann 等ICML 2025
- COMPASS: SRAM-Based Computing-in-Memory SNN Accelerator with Adaptive Spike SpeculationZongwu Wang, Fangxin Liu, Ning Yang, Shiyuan Huang 等MICRO 2024 · 被引用 11 次
