Operation-Aware Soft Channel Pruning using Differentiable Masks
Minsoo Kang, Bohyung Han
摘要
We propose a simple but effective data-driven channel pruning algorithm, which compresses deep neural networks in a differentiable way by exploiting the characteristics of operations. The proposed approach makes a joint consideration of batch normalization (BN) and rectified linear unit (ReLU) for channel pruning; it estimates how likely the two successive operations deactivate each feature map and prunes the channels with high probabilities. To this end, we learn differentiable masks for individual channels and make soft decisions throughout the optimization procedure, which facilitates to explore larger search space and train more stable networks. The proposed framework enables us to identify compressed models via a joint learning of model parameters and channel pruning without an extra procedure of fine-tuning. We perform extensive experiments and achieve outstanding performance in terms of the accuracy of output networks given the same amount of resources when compared with the state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper36
- IRPruneDet: Efficient Infrared Small Target Detection via Wavelet Structure-Regularized Soft Channel PruningMingjin Zhang, Handi Yang, Jie Guo, Yunsong Li 等AAAI 2024 · 被引用 159 次
- Only Train Once: A One-Shot Neural Network Training And Pruning FrameworkTianyi Chen, Bo Ji, Tianyu Ding, Biyi Fang 等NeurIPS 2021 · 被引用 135 次
- DiffRate : Differentiable Compression Rate for Efficient Vision TransformersMengzhao Chen, Wenqi Shao, Peng Xu, Mingbao Lin 等ICCV 2023 · 被引用 87 次
- CHEX: CHannel EXploration for CNN Model CompressionZejiang Hou, Minghai Qin, Fei Sun, Xiaolong Ma 等CVPR 2022 · 被引用 80 次
- Aligned Structured Sparsity Learning for Efficient Image Super-ResolutionYulun Zhang, Huan Wang, Can Qin, Yun FuNeurIPS 2021 · 被引用 72 次
它引用的顶会 Paper1
相关 Paper
- Automatic Channel Pruning with Hyper-parameter Search and Dynamic MaskingBaopu Li, Yanwen Fan, Zhihong Pan, Yuchen Bian 等ACM MM 2021 · 被引用 3 次
- Exploring Gradient Flow Based Saliency for DNN Model CompressionXinyu Liu, Baopu Li, Zhen Chen, Yixuan YuanACM MM 2021 · 被引用 7 次
- SepPrune: Structured Pruning for Efficient Deep Speech SeparationYuqi Li, Kai Li, Xin Yin, Zhifei Yang 等AAAI 2026 · 被引用 4 次
- Storage Efficient and Dynamic Flexible Runtime Channel Pruning via Deep Reinforcement LearningJianda Chen, Shangyu Chen, Sinno Jialin PanNeurIPS 2020 · 被引用 31 次
- Bayesian based Re-parameterization for DNN Model PruningXiaotong Lu, Teng Xi, Baopu Li, Gang Zhang 等ACM MM 2022 · 被引用 4 次
