Plug-in, Trainable Gate for Streamlining Arbitrary Neural Networks
Jaedeok Kim, Chiyoun Park, Hyun-Joo Jung, Yoonsuck Choe
Abstract
Architecture optimization, which is a technique for finding an efficient neural network that meets certain requirements, generally reduces to a set of multiple-choice selection problems among alternative sub-structures or parameters. The discrete nature of the selection problem, however, makes this optimization difficult. To tackle this problem we introduce a novel concept of a trainable gate function. The trainable gate function, which confers a differentiable property to discrete-valued variables, allows us to directly optimize loss functions that include non-differentiable discrete values such as 0-1 selection. The proposed trainable gate can be applied to pruning. Pruning can be carried out simply by appending the proposed trainable gate functions to each intermediate output tensor followed by fine-tuning the overall model, using any gradient-based training methods. So the proposed method can jointly optimize the selection of the pruned channels while fine-tuning the weights of the pruned model at the same time. Our experimental results demonstrate that the proposed method efficiently optimizes arbitrary neural networks in various tasks such as image classification, style transfer, optical flow estimation, and neural machine translation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4c9ff026-a009-450e-8a3e-b309d76ae7aaCited by top-tier papers7
- Injecting Logical Constraints into Neural Networks via Straight-Through EstimatorsZhun Yang, Joohyung Lee, Chiyoun ParkICML 2022 · 26 citations
- Structural Alignment for Network Pruning through Partial RegularizationShangqian Gao, Zeyu Zhang, Yanfu Zhang, Feihu Huang et al.ICCV 2023 · 26 citations
- Auto- Train-Once: Controller Network Guided Automatic Network Pruning from ScratchXidong Wu, Shangqian Gao, Zeyu Zhang, Zhenzhen Li et al.CVPR 2024 · 13 citations
- Pruning Parameterization with Bi-level Optimization for Efficient Semantic Segmentation on the EdgeChangdi Yang, Pu Zhao, Yanyu Li, Wei Niu et al.CVPR 2023
- Discrete Model Compression With Resource Constraint for Deep Neural NetworksShangqian Gao, Feihu Huang, Jian Pei, Heng HuangCVPR 2020
Related papers
- Deep Differentiable Logic Gate NetworksFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenNeurIPS 2022 · 117 citations
- DMCP: Differentiable Markov Channel Pruning for Neural NetworksShaopeng Guo, Yujie Wang, Quanquan Li, Junjie YanCVPR 2020
- S2HPruner: Soft-to-Hard Distillation Bridges the Discretization Gap in PruningWeihao Lin, Shengji Tang, Chong Yu, Peng Ye et al.NeurIPS 2024 · 2 citations
- Deterministic Differentiable Structured Pruning for Large Language ModelsWeiyu Huang, Pengle Zhang, Xiaolu Zhang, JUN ZHOU et al.ICML 2026 · 1 citation
- GDP: Stabilized Neural Network Pruning via Gates with Differentiable PolarizationYi Guo, Huan Yuan, Jianchao Tan, Zhangyang Wang et al.ICCV 2021 · 52 citations
