ReSprop: Reuse Sparsified Backpropagation
Negar Goli, Tor M. Aamodt
Abstract
The success of Convolutional Neural Networks (CNNs) in various applications is accompanied by a significant increase in computation and training time. In this work, we focus on accelerating training by observing that about 90% of gradients are reusable during training. Leveraging this observation, we propose a new algorithm, Reuse-Sparse-Backprop (ReSprop), as a method to sparsify gradient vectors during CNN training. ReSprop maintains stateof-the-art accuracy on CIFAR-10, CIFAR-100, and Ima-geNet datasets with less than 1.1% accuracy loss while enabling a reduction in back-propagation computations by a factor of 10× resulting in a 2.7× overall speedup in training. As the computation reduction introduced by Re-Sprop is accomplished by introducing fine-grained sparsity that reduces computation efficiency on GPUs, we introduce a generic sparse convolution neural network accelerator (GSCN), which is designed to accelerate sparse backpropagation convolutions. When combined with ReSprop, GSCN achieves 8.0× and 7.2× speedup in the backward pass on ResNet34 and VGG16 versus a GTX 1080 Ti GPU.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c7705937-a3ff-4cb4-9103-d4602c109bb2Cited by top-tier papers9
- Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network ModelsBeidi Chen, Tri Dao, Kaizhao Liang, Jiaming Yang et al.ICLR 2022 · 94 citations
- ZeroFL: Efficient On-Device Training for Federated Learning with Local SparsityXinchi Qiu, Javier Fernández-Marqués, Pedro P. B. de Gusmao, Yan Gao et al.ICLR 2022 · 87 citations
- Sparse Weight Activation TrainingMd Aamir Raihan, Tor M. AamodtNeurIPS 2020 · 83 citations
- Faster Neural Network Training with Approximate Tensor OperationsMenachem Adelman, Kfir Y. Levy, Ido Hakimi, Mark SilbersteinNeurIPS 2021 · 30 citations
- Anticipating and eliminating redundant computations in accelerated sparse trainingJonathan S. Lew, Yunpeng Liu, Wenyi Gong, Negar Goli et al.ISCA 2022 · 10 citations
Related papers
- SparseTrain: Exploiting Dataflow Sparsity for Efficient Convolutional Neural Networks TrainingPengcheng Dai, Jianlei Yang, Xucheng Ye, Xingzhou Cheng et al.DAC 2020 · 27 citations
- ADA-GP: Accelerating DNN Training By Adaptive Gradient PredictionVahid Janfaza, Shantanu Mandal, Farabi Mahmud, Abdullah MuzahidMICRO 2023 · 3 citations
- RSC: Accelerate Graph Neural Networks Training via Randomized Sparse ComputationsZirui Liu, Shengyuan Chen, Kaixiong Zhou, Daochen Zha et al.ICML 2023 · 24 citations
- SparseProp: Efficient Sparse Backpropagation for Faster Training of Neural Networks at the EdgeMahdi Nikdan, Tommaso Pegolotti, Eugenia Iofinova, Eldar Kurtic et al.ICML 2023 · 14 citations
- An In-depth Study of Stochastic BackpropagationJun Fang, Mingze Xu, Hao Chen, Bing Shuai et al.NeurIPS 2022 · 2 citations
