Differentiable Transportation Pruning
Yunqiang Li, Jan C. van Gemert, Torsten Hoefler, Bert Moons, Evangelos Eleftheriou, Bram-Ernst Verhoef
Abstract
Deep learning algorithms are increasingly employed at the edge. However, edge devices are resource constrained and thus require efficient deployment of deep neural networks. Pruning methods are a key tool for edge deployment as they can improve storage, compute, memory bandwidth, and energy usage. In this paper we propose a novel accurate pruning technique that allows precise control over the output network size. Our method uses an efficient optimal transportation scheme which we make end-to-end differentiable and which automatically tunes the exploration-exploitation behavior of the algorithm to find accurate sparse sub-networks. We show that our method achieves state-of-the-art performance compared to previous pruning methods on 3 different datasets, using 5 different models, across a wide range of pruning ratios, and with two types of sparsity budgets and pruning granularities.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Auto- Train-Once: Controller Network Guided Automatic Network Pruning from ScratchXidong Wu, Shangqian Gao, Zeyu Zhang, Zhenzhen Li et al.CVPR 2024 · 13 citations
- REPrune: Channel Pruning via Kernel Representative SelectionMincheol Park, Dongjin Kim, Cheonjun Park, Yuna Park et al.AAAI 2024 · 5 citations
- S2HPruner: Soft-to-Hard Distillation Bridges the Discretization Gap in PruningWeihao Lin, Shengji Tang, Chong Yu, Peng Ye et al.NeurIPS 2024 · 2 citations
- BilevelPruning: Unified Dynamic and Static Channel Pruning for Convolutional Neural NetworksShangqian Gao, Yanfu Zhang, Feihu Huang, Heng HuangCVPR 2024
- MDP: Multidimensional Vision Model Pruning with Latency ConstraintXinglong Sun, Barath Lakshmanan, Maying Shen, Shiyi Lan et al.CVPR 2025
Builds on24
- MetaPruning: Meta Learning for Automatic Neural Network Channel PruningZechun Liu, Haoyuan Mu, Xiangyu Zhang, Zichao Guo et al.ICCV 2019 · 633 citations
- Dynamic Model Pruning with FeedbackTao Lin, Sebastian U. Stich, Luis Barba, Daniil Dmitriev et al.ICLR 2020 · 229 citations
- WoodFisher: Efficient Second-Order Approximation for Neural Network CompressionSidak Pal Singh, Dan AlistarhNeurIPS 2020 · 217 citations
- Neural Pruning via Growing RegularizationHuan Wang, Can Qin, Yulun Zhang, Yun FuICLR 2021 · 188 citations
- Operation-Aware Soft Channel Pruning using Differentiable MasksMinsoo Kang, Bohyung HanICML 2020 · 165 citations
Related papers
- Device-Wise Federated Network PruningShangqian Gao, Junyi Li, Zeyu Zhang, Yanfu Zhang et al.CVPR 2024
- Rethinking Pruning for Accelerating Deep Inference At the EdgeDawei Gao, Xiaoxi He, Zimu Zhou, Yongxin Tong et al.KDD 2020 · 24 citations
- PowerPruning: Selecting Weights and Activations for Power-Efficient Neural Network AccelerationRichard Petri, Grace Li Zhang, Yiran Chen, Ulf Schlichtmann et al.DAC 2023 · 11 citations
- Automatic Channel Pruning with Hyper-parameter Search and Dynamic MaskingBaopu Li, Yanwen Fan, Zhihong Pan, Yuchen Bian et al.ACM MM 2021 · 3 citations
- Towards Higher Ranks via Adversarial Weight PruningYuchuan Tian, Hanting Chen, Tianyu Guo, Chao Xu et al.NeurIPS 2023 · 9 citations
