Operation-Aware Soft Channel Pruning using Differentiable Masks
Minsoo Kang, Bohyung Han
Abstract
We propose a simple but effective data-driven channel pruning algorithm, which compresses deep neural networks in a differentiable way by exploiting the characteristics of operations. The proposed approach makes a joint consideration of batch normalization (BN) and rectified linear unit (ReLU) for channel pruning; it estimates how likely the two successive operations deactivate each feature map and prunes the channels with high probabilities. To this end, we learn differentiable masks for individual channels and make soft decisions throughout the optimization procedure, which facilitates to explore larger search space and train more stable networks. The proposed framework enables us to identify compressed models via a joint learning of model parameters and channel pruning without an extra procedure of fine-tuning. We perform extensive experiments and achieve outstanding performance in terms of the accuracy of output networks given the same amount of resources when compared with the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c26fc63a-7dae-4575-afbc-5072ae1a8d47Cited by top-tier papers36
- IRPruneDet: Efficient Infrared Small Target Detection via Wavelet Structure-Regularized Soft Channel PruningMingjin Zhang, Handi Yang, Jie Guo, Yunsong Li et al.AAAI 2024 · 159 citations
- Only Train Once: A One-Shot Neural Network Training And Pruning FrameworkTianyi Chen, Bo Ji, Tianyu Ding, Biyi Fang et al.NeurIPS 2021 · 135 citations
- DiffRate : Differentiable Compression Rate for Efficient Vision TransformersMengzhao Chen, Wenqi Shao, Peng Xu, Mingbao Lin et al.ICCV 2023 · 87 citations
- CHEX: CHannel EXploration for CNN Model CompressionZejiang Hou, Minghai Qin, Fei Sun, Xiaolong Ma et al.CVPR 2022 · 80 citations
- Aligned Structured Sparsity Learning for Efficient Image Super-ResolutionYulun Zhang, Huan Wang, Can Qin, Yun FuNeurIPS 2021 · 72 citations
Builds on1
Related papers
- Automatic Channel Pruning with Hyper-parameter Search and Dynamic MaskingBaopu Li, Yanwen Fan, Zhihong Pan, Yuchen Bian et al.ACM MM 2021 · 3 citations
- Exploring Gradient Flow Based Saliency for DNN Model CompressionXinyu Liu, Baopu Li, Zhen Chen, Yixuan YuanACM MM 2021 · 7 citations
- SepPrune: Structured Pruning for Efficient Deep Speech SeparationYuqi Li, Kai Li, Xin Yin, Zhifei Yang et al.AAAI 2026 · 4 citations
- Storage Efficient and Dynamic Flexible Runtime Channel Pruning via Deep Reinforcement LearningJianda Chen, Shangyu Chen, Sinno Jialin PanNeurIPS 2020 · 31 citations
- Bayesian based Re-parameterization for DNN Model PruningXiaotong Lu, Teng Xi, Baopu Li, Gang Zhang et al.ACM MM 2022 · 4 citations
