Channel Regeneration: Improving Channel Utilization for Compact DNNs
Ankit Kumar Sharma, Hassan Foroosh
Abstract
Overparameterized deep neural networks have redundant neurons that do not contribute to the network's accuracy. In this paper, we introduce a novel channel regeneration technique that reinvigorates these redundant channels of efficient architectures by re-initializing its batch normalization scaling factor γ. This re-initialization of BN γ of these channels promotes regular weight updates during training. Furthermore, we show that channel regeneration encourages the channels to contribute equally to the learned representation and further boosts the generalization accuracy. We apply our technique at regular intervals of the training cycle to improve channel utilization. The solutions proposed in previous works either raise the total computational cost or increase the model complexity. Integrating the channel regeneration technique into the training methodology of efficient architectures requires minimal effort and comes at no additional cost in size or memory. Extensive experiments on several image classification benchmarks and on semantic segmentation task demonstrate the effectiveness of applying the channel regeneration technique to compact architectures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5667d6e-1552-47a4-a028-89f108a8ca1dBuilds on5
- Drawing Early-Bird Tickets: Toward More Efficient Training of Deep NetworksHaoran You, Chaojian Li, Pengfei Xu, Yonggan Fu et al.ICLR 2020 · 282 citations
- Channel Equilibrium Networks for Learning Deep RepresentationWenqi Shao, Shitao Tang, Xingang Pan, Ping Tan et al.ICML 2020 · 17 citations
- Learning Filter Pruning Criteria for Deep Convolutional Neural Networks AccelerationYang He, Yuhang Ding, Ping Liu, Linchao Zhu et al.CVPR 2020
- HRank: Filter Pruning Using High-Rank Feature MapMingbao Lin, Rongrong Ji, Yan Wang, Yichen Zhang et al.CVPR 2020
- Group Sparsity: The Hinge Between Filter Pruning and Decomposition for Network CompressionYawei Li, Shuhang Gu, Christoph Mayer, Luc Van Gool et al.CVPR 2020
Related papers
- Deconstructing the Regularization of BatchNormYann N. Dauphin, Ekin Dogus CubukICLR 2021 · 6 citations
- InterBN: Channel Fusion for Adversarial Unsupervised Domain AdaptationMengzhu Wang, Wei Wang, Baopu Li, Xiang Zhang et al.ACM MM 2021 · 27 citations
- Batch Normalization Biases Residual Blocks Towards the Identity Function in Deep NetworksSoham De, Samuel L. SmithNeurIPS 2020 · 173 citations
- Exploring Gradient Flow Based Saliency for DNN Model CompressionXinyu Liu, Baopu Li, Zhen Chen, Yixuan YuanACM MM 2021 · 7 citations
- Operation-Aware Soft Channel Pruning using Differentiable MasksMinsoo Kang, Bohyung HanICML 2020 · 165 citations
