Can we get the best of both Binary Neural Networks and Spiking Neural Networks for Efficient Computer Vision?
Gourav Datta, Zeyu Liu, Peter Anthony Beerel
摘要
Binary Neural networks (BNN) have emerged as an attractive computing paradigm for a wide range of low-power vision tasks. However, state-of-theart (SOTA) BNNs do not yield any sparsity, and induce a significant number of non-binary operations. On the other hand, activation sparsity can be provided by spiking neural networks (SNN), that too have gained significant traction in recent times. Thanks to this sparsity, SNNs when implemented on neuromorphic hardware, have the potential to be significantly more power-efficient compared to traditional artifical neural networks (ANN). However, SNNs incur multiple time steps to achieve close to SOTA accuracy. Ironically, this increases latency and energy-costs that SNNs were proposed to reduce-and presents itself as a major hurdle in realizing SNNs' theoretical gains in practice. This raises an intriguing question: Can we obtain SNN-like sparsity and BNN-like accuracy and enjoy the energy-efficiency benefits of both? To answer this question, in this paper, we present a training framework for sparse binary activation neural networks (BANN) using a novel variant of the Hoyer regularizer. We estimate the threshold of each BANN layer as the Hoyer extremum of a clipped version of its activation map, where the clipping value is trained using gradient descent with our Hoyer regularizer. This approach shifts the activation values away from the threshold, thereby mitigating the effect of noise that can otherwise degrade the BANN accuracy. Our approach outperforms existing BNNs, SNNs, and adder neural networks (that also avoid energy-expensive multiplication operations similar to BNNs and SNNs) in terms of the accuracy-FLOPs trade-off for complex image recognition tasks. Downstream experiments on object detection further demonstrate the efficacy of our approach. Lastly, we demonstrate the portability of our approach to SNNs with multiple time steps. Codes are publicly available here. However, most of these efforts induce a significant number of non-binary operations which degrade the computational efficiency. For example, ReactNet-based BNNs (Liu et al., 2020a; Zhijun Tu & Wang, 2022b) incur custom non-linear functions, including RPReLU that are significantly more complex compared to threshold or ReLU operations, duplicated basic blocks that significantly increase the total number of floating point operations (FLOPs) and parameter count. Moreover, all . This extremum is the minimum, because the second derivative is greater than zero for any value of the output element. Training with the Hoyer regularizer can effectively help push the activation values that are larger than the extremum (u l >E(u l )) even larger and those that are smaller than the extremum (u l <E(u l )) even smaller.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper26
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang 等NeurIPS 2021 · 被引用 857 次
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu 等AAAI 2021 · 被引用 694 次
- Spiking-YOLO: Spiking Neural Network for Energy-Efficient Object DetectionSei Joon Kim, Seongsik Park, Byunggook Na, Sungroh YoonAAAI 2020 · 被引用 512 次
- Temporal Efficient Training of Spiking Neural Network via Gradient Re-weightingShikuang Deng, Yuhang Li, Shanghang Zhang, Shi GuICLR 2022 · 被引用 361 次
- Enabling Deep Spiking Neural Networks with Hybrid Conversion and Spike Timing Dependent BackpropagationNitin Rathi, Gopalakrishnan Srinivasan, Priyadarshini Panda, Kaushik RoyICLR 2020 · 被引用 347 次
相关 Paper
- Activity Pruning for Efficient Spiking Neural NetworksTong Bu, Xinyu Shi, Zhaofei YuNeurIPS 2025 · 被引用 2 次
- SpikeConverter: An Efficient Conversion Framework Zipping the Gap between Artificial Neural Networks and Spiking Neural NetworksFangxin Liu, Wenbo Zhao, Yongbiao Chen, Zongwu Wang 等AAAI 2022 · 被引用 50 次
- A Unified Optimization Framework of ANN-SNN Conversion: Towards Optimal Mapping from Activation Values to Firing RatesHaiyan Jiang, Srinivas Anumasa, Giulia De Masi, Huan Xiong 等ICML 2023 · 被引用 37 次
- Skipper: Enabling efficient SNN training through activation-checkpointing and time-skippingSonali Singh, Anup Sarma, Sen Lu, Abhronil Sengupta 等MICRO 2022 · 被引用 13 次
- Are Conventional SNNs Really Efficient? A Perspective from Network QuantizationGuobin Shen, Dongcheng Zhao, Tenglong Li, Jindong Li 等CVPR 2024 · 被引用 8 次
