Control Variate Approximation for DNN Accelerators
Georgios Zervakis, Ourania Spantidi, Iraklis Anagnostopoulos, Hussam Amrouch, Jörg Henkel
摘要
In this work, we introduce a control variate approximation technique for low error approximate Deep Neural Network (DNN) accelerators. The control variate technique is used in Monte Carlo methods to achieve variance reduction. Our approach significantly decreases the induced error due to approximate multiplications in DNN inference, without requiring time-exhaustive retraining compared to state-of-the-art. Leveraging our control variate method, we use highly approximated multipliers to generate power-optimized DNN accelerators. Our experimental evaluation on six DNNs, for Cifar-10 and Cifar-100 datasets, demonstrates that, compared to the accurate design, our control variate approximation achieves same performance and 24% power reduction for a merely 0.16% accuracy loss.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- PowerPruning: Selecting Weights and Activations for Power-Efficient Neural Network AccelerationRichard Petri, Grace Li Zhang, Yiran Chen, Ulf Schlichtmann 等DAC 2023 · 被引用 11 次
- Neural Control Variates with Automatic IntegrationZilu Li, Guandao Yang, Qingqing Zhao, Xi Deng 等SIGGRAPH 2024 · 被引用 9 次
- Arbitor: A Numerically Accurate Hardware Emulation Tool for DNN AcceleratorsChenhao Jiang, Anand Jayarajan, Hao Lu, Gennady PekhimenkoUSENIX ATC 2023 · 被引用 5 次
- Faster Neural Network Training with Approximate Tensor OperationsMenachem Adelman, Kfir Y. Levy, Ido Hakimi, Mark SilbersteinNeurIPS 2021 · 被引用 30 次
- COSAIM: Counter-based Stochastic-behaving Approximate Integer Multiplier for Deep Neural NetworksShuyuan Yu, Yibo Liu, Sheldon X.-D. TanDAC 2021 · 被引用 14 次
