NeuralFuse: Learning to Recover the Accuracy of Access-Limited Neural Network Inference in Low-Voltage Regimes
Hao-Lun Sun, Lei Hsiung, Nandhini Chandramoorthy, Pin-Yu Chen, Tsung-Yi Ho
摘要
Deep neural networks (DNNs) have become ubiquitous in machine learning, but their energy consumption remains problematically high. An effective strategy for reducing such consumption is supply-voltage reduction, but if done too aggressively, it can lead to accuracy degradation. This is due to random bit-flips in static random access memory (SRAM), where model parameters are stored. To address this challenge, we have developed NeuralFuse, a novel add-on module that handles the energy-accuracy tradeoff in low-voltage regimes by learning input transformations and using them to generate error-resistant data representations, thereby protecting DNN accuracy in both nominal and low-voltage scenarios. As well as being easy to implement, NeuralFuse can be readily applied to DNNs with limited access, such cloud-based APIs that are accessed remotely or non-configurable hardware. Our experimental results demonstrate that, at a 1% bit-error rate, NeuralFuse can reduce SRAM access energy by up to 24% while recovering accuracy by up to 57%. To the best of our knowledge, this is the first approach to addressing low-voltage-induced bit errors that requires no model retraining.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
相关 Paper
- SparkXD: A Framework for Resilient and Energy-Efficient Spiking Neural Network Inference using Approximate DRAMRachmad Vidya Wicaksana Putra, Muhammad Abdullah Hanif, Muhammad ShafiqueDAC 2021 · 被引用 28 次
- Fault-free: A Fault-resilient Deep Neural Network Accelerator based on Realistic ReRAM DevicesHyein Shin, Myeonggu Kang, Lee-Sup KimDAC 2021 · 被引用 19 次
- Terminal Brain Damage: Exposing the Graceless Degradation in Deep Neural Networks Under Hardware Fault AttacksSanghyun Hong, Pietro Frigo, Yigitcan Kaya, Cristiano Giuffrida 等USENIX Security 2019 · 被引用 255 次
- Boosting Deep Neural Network Efficiency with Dual-Module InferenceLiu Liu, Lei Deng, Zhaodong Chen, Yuke Wang 等ICML 2020 · 被引用 9 次
- Bipolar vector classifier for fault-tolerant deep neural networksSuyong Lee, Insu Choi, Joon-Sung YangDAC 2022 · 被引用 4 次
