Control Variate Approximation for DNN Accelerators
Georgios Zervakis, Ourania Spantidi, Iraklis Anagnostopoulos, Hussam Amrouch, Jörg Henkel
Abstract
In this work, we introduce a control variate approximation technique for low error approximate Deep Neural Network (DNN) accelerators. The control variate technique is used in Monte Carlo methods to achieve variance reduction. Our approach significantly decreases the induced error due to approximate multiplications in DNN inference, without requiring time-exhaustive retraining compared to state-of-the-art. Leveraging our control variate method, we use highly approximated multipliers to generate power-optimized DNN accelerators. Our experimental evaluation on six DNNs, for Cifar-10 and Cifar-100 datasets, demonstrates that, compared to the accurate design, our control variate approximation achieves same performance and 24% power reduction for a merely 0.16% accuracy loss.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- PowerPruning: Selecting Weights and Activations for Power-Efficient Neural Network AccelerationRichard Petri, Grace Li Zhang, Yiran Chen, Ulf Schlichtmann et al.DAC 2023 · 11 citations
- Neural Control Variates with Automatic IntegrationZilu Li, Guandao Yang, Qingqing Zhao, Xi Deng et al.SIGGRAPH 2024 · 9 citations
- Arbitor: A Numerically Accurate Hardware Emulation Tool for DNN AcceleratorsChenhao Jiang, Anand Jayarajan, Hao Lu, Gennady PekhimenkoUSENIX ATC 2023 · 5 citations
- Faster Neural Network Training with Approximate Tensor OperationsMenachem Adelman, Kfir Y. Levy, Ido Hakimi, Mark SilbersteinNeurIPS 2021 · 30 citations
- COSAIM: Counter-based Stochastic-behaving Approximate Integer Multiplier for Deep Neural NetworksShuyuan Yu, Yibo Liu, Sheldon X.-D. TanDAC 2021 · 14 citations
