PowerPruning: Selecting Weights and Activations for Power-Efficient Neural Network Acceleration
Richard Petri, Grace Li Zhang, Yiran Chen, Ulf Schlichtmann, Bing Li
摘要
Deep neural networks (DNNs) have been successfully applied in various fields. A major challenge of deploying DNNs, especially on edge devices, is power consumption, due to the large number of multiply-and-accumulate (MAC) operations. To address this challenge, we propose PowerPruning, a novel method to reduce power consumption in digital neural network accelerators by selecting weights that lead to less power consumption in MAC operations. In addition, the timing characteristics of the selected weights together with all activation transitions are evaluated. The weights and activations that lead to small delays are further selected. Consequently, the maximum delay of the sensitized circuit paths in the MAC units is reduced even without modifying MAC units, which thus allows a flexible scaling of supply voltage to reduce power consumption further. Together with retraining, the proposed method can reduce power consumption of DNNs on hardware by up to 73.9% with only a slight accuracy loss.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Bit-Pruning: A Sparse Multiplication-Less Dot-ProductYusuke Sekikawa, Shingo YashimaICLR 2023
- Control Variate Approximation for DNN AcceleratorsGeorgios Zervakis, Ourania Spantidi, Iraklis Anagnostopoulos, Hussam Amrouch 等DAC 2021 · 被引用 32 次
- PIM-Prune: Fine-Grain DCNN Pruning for Crossbar-Based Process-In-Memory ArchitectureChaoqun Chu, Yanzhi Wang, Yilong Zhao, Xiaolong Ma 等DAC 2020 · 被引用 64 次
- Intermittent-Aware Neural Network PruningChih-Chia Lin, Chia-Yin Liu, Chih-Hsuan Yen, Tei-Wei Kuo 等DAC 2023 · 被引用 11 次
- DPACS: Hardware Accelerated Dynamic Neural Network Pruning through Algorithm-Architecture Co-designYizhao Gao, Baoheng Zhang, Xiaojuan Qi, Hayden Kwok-Hay SoASPLOS 2023 · 被引用 13 次
