Network Pruning via Performance Maximization
Shangqian Gao, Feihu Huang, Weidong Cai, Heng Huang
Abstract
Channel pruning is a class of powerful methods for model compression. When pruning a neural network, it's ideal to obtain a sub-network with higher accuracy. However, a sub-network does not necessarily have high accuracy with low classification loss (loss-metric mismatch). In the paper, we first consider the loss-metric mismatch problem for pruning and propose a novel channel pruning method for Convolutional Neural Networks (CNNs) by directly maximizing the performance (i.e., accuracy) of subnetworks. Specifically, we train a stand-alone neural network to predict sub-networks' performance and then maximize the output of the network as a proxy of accuracy to guide pruning. Training such a performance prediction network efficiently is not an easy task, and it may potentially suffer from the problem of catastrophic forgetting and the imbalance distribution of sub-networks. To deal with this challenge, we introduce a corresponding episodic memory to update and collect sub-networks during the pruning process. In the experiment section, we further demonstrate that the gradients from the performance prediction network and the classification loss have different directions. Extensive experimental results show that the proposed method can achieve state-of-the-art performance with ResNet,
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext da2a95af-7ac6-47f8-a7bf-378995289f46Cited by top-tier papers20
- SAViT: Structure-Aware Vision Transformer Pruning via Collaborative OptimizationChuanyang Zheng, Zheyang Li, Kai Zhang, Zhi Yang et al.NeurIPS 2022 · 98 citations
- CHEX: CHannel EXploration for CNN Model CompressionZejiang Hou, Minghai Qin, Fei Sun, Xiaolong Ma et al.CVPR 2022 · 80 citations
- Topology-Aware Network Pruning using Multi-stage Graph Embedding and Reinforcement LearningSixing Yu, Arya Mazaheri, Ali JannesariICML 2022 · 54 citations
- Automatic Network Pruning via Hilbert-Schmidt Independence Criterion Lasso under Information Bottleneck PrincipleSong Guo, Lei Zhang, Xiawu Zheng, Yan Wang et al.ICCV 2023 · 30 citations
- Structural Alignment for Network Pruning through Partial RegularizationShangqian Gao, Zeyu Zhang, Yanfu Zhang, Feihu Huang et al.ICCV 2023 · 26 citations
Builds on10
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- MetaPruning: Meta Learning for Automatic Neural Network Channel PruningZechun Liu, Haoyuan Mu, Xiangyu Zhang, Zichao Guo et al.ICCV 2019 · 633 citations
- Operation-Aware Soft Channel Pruning using Differentiable MasksMinsoo Kang, Bohyung HanICML 2020 · 165 citations
- Good Subnetworks Provably Exist: Pruning via Greedy Forward SelectionMao Ye, Chengyue Gong, Lizhen Nie, Denny Zhou et al.ICML 2020 · 123 citations
- Plug-in, Trainable Gate for Streamlining Arbitrary Neural NetworksJaedeok Kim, Chiyoun Park, Hyun-Joo Jung, Yoonsuck ChoeAAAI 2020 · 19 citations
Related papers
- ResRep: Lossless CNN Pruning via Decoupling Remembering and ForgettingXiaohan Ding, Tianxiang Hao, Jianchao Tan, Ji Liu et al.ICCV 2021 · 202 citations
- Device-Wise Federated Network PruningShangqian Gao, Junyi Li, Zeyu Zhang, Yanfu Zhang et al.CVPR 2024
- BilevelPruning: Unified Dynamic and Static Channel Pruning for Convolutional Neural NetworksShangqian Gao, Yanfu Zhang, Feihu Huang, Heng HuangCVPR 2024
- DMCP: Differentiable Markov Channel Pruning for Neural NetworksShaopeng Guo, Yujie Wang, Quanquan Li, Junjie YanCVPR 2020
- Linearly Replaceable Filters for Deep Network Channel PruningDonggyu Joo, Eojindl Yi, Sunghyun Baek, Junmo KimAAAI 2021 · 39 citations
