Channel Regeneration: Improving Channel Utilization for Compact DNNs
Ankit Kumar Sharma, Hassan Foroosh
摘要
Overparameterized deep neural networks have redundant neurons that do not contribute to the network's accuracy. In this paper, we introduce a novel channel regeneration technique that reinvigorates these redundant channels of efficient architectures by re-initializing its batch normalization scaling factor γ. This re-initialization of BN γ of these channels promotes regular weight updates during training. Furthermore, we show that channel regeneration encourages the channels to contribute equally to the learned representation and further boosts the generalization accuracy. We apply our technique at regular intervals of the training cycle to improve channel utilization. The solutions proposed in previous works either raise the total computational cost or increase the model complexity. Integrating the channel regeneration technique into the training methodology of efficient architectures requires minimal effort and comes at no additional cost in size or memory. Extensive experiments on several image classification benchmarks and on semantic segmentation task demonstrate the effectiveness of applying the channel regeneration technique to compact architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Drawing Early-Bird Tickets: Toward More Efficient Training of Deep NetworksHaoran You, Chaojian Li, Pengfei Xu, Yonggan Fu 等ICLR 2020 · 被引用 282 次
- Channel Equilibrium Networks for Learning Deep RepresentationWenqi Shao, Shitao Tang, Xingang Pan, Ping Tan 等ICML 2020 · 被引用 17 次
- Learning Filter Pruning Criteria for Deep Convolutional Neural Networks AccelerationYang He, Yuhang Ding, Ping Liu, Linchao Zhu 等CVPR 2020
- HRank: Filter Pruning Using High-Rank Feature MapMingbao Lin, Rongrong Ji, Yan Wang, Yichen Zhang 等CVPR 2020
- Group Sparsity: The Hinge Between Filter Pruning and Decomposition for Network CompressionYawei Li, Shuhang Gu, Christoph Mayer, Luc Van Gool 等CVPR 2020
相关 Paper
- Deconstructing the Regularization of BatchNormYann N. Dauphin, Ekin Dogus CubukICLR 2021 · 被引用 6 次
- InterBN: Channel Fusion for Adversarial Unsupervised Domain AdaptationMengzhu Wang, Wei Wang, Baopu Li, Xiang Zhang 等ACM MM 2021 · 被引用 27 次
- Batch Normalization Biases Residual Blocks Towards the Identity Function in Deep NetworksSoham De, Samuel L. SmithNeurIPS 2020 · 被引用 173 次
- Exploring Gradient Flow Based Saliency for DNN Model CompressionXinyu Liu, Baopu Li, Zhen Chen, Yixuan YuanACM MM 2021 · 被引用 7 次
- Operation-Aware Soft Channel Pruning using Differentiable MasksMinsoo Kang, Bohyung HanICML 2020 · 被引用 165 次
