Plug-in, Trainable Gate for Streamlining Arbitrary Neural Networks
Jaedeok Kim, Chiyoun Park, Hyun-Joo Jung, Yoonsuck Choe
摘要
Architecture optimization, which is a technique for finding an efficient neural network that meets certain requirements, generally reduces to a set of multiple-choice selection problems among alternative sub-structures or parameters. The discrete nature of the selection problem, however, makes this optimization difficult. To tackle this problem we introduce a novel concept of a trainable gate function. The trainable gate function, which confers a differentiable property to discrete-valued variables, allows us to directly optimize loss functions that include non-differentiable discrete values such as 0-1 selection. The proposed trainable gate can be applied to pruning. Pruning can be carried out simply by appending the proposed trainable gate functions to each intermediate output tensor followed by fine-tuning the overall model, using any gradient-based training methods. So the proposed method can jointly optimize the selection of the pruned channels while fine-tuning the weights of the pruned model at the same time. Our experimental results demonstrate that the proposed method efficiently optimizes arbitrary neural networks in various tasks such as image classification, style transfer, optical flow estimation, and neural machine translation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Injecting Logical Constraints into Neural Networks via Straight-Through EstimatorsZhun Yang, Joohyung Lee, Chiyoun ParkICML 2022 · 被引用 26 次
- Structural Alignment for Network Pruning through Partial RegularizationShangqian Gao, Zeyu Zhang, Yanfu Zhang, Feihu Huang 等ICCV 2023 · 被引用 26 次
- Auto- Train-Once: Controller Network Guided Automatic Network Pruning from ScratchXidong Wu, Shangqian Gao, Zeyu Zhang, Zhenzhen Li 等CVPR 2024 · 被引用 13 次
- Pruning Parameterization with Bi-level Optimization for Efficient Semantic Segmentation on the EdgeChangdi Yang, Pu Zhao, Yanyu Li, Wei Niu 等CVPR 2023
- Discrete Model Compression With Resource Constraint for Deep Neural NetworksShangqian Gao, Feihu Huang, Jian Pei, Heng HuangCVPR 2020
相关 Paper
- Deep Differentiable Logic Gate NetworksFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenNeurIPS 2022 · 被引用 117 次
- DMCP: Differentiable Markov Channel Pruning for Neural NetworksShaopeng Guo, Yujie Wang, Quanquan Li, Junjie YanCVPR 2020
- S2HPruner: Soft-to-Hard Distillation Bridges the Discretization Gap in PruningWeihao Lin, Shengji Tang, Chong Yu, Peng Ye 等NeurIPS 2024 · 被引用 2 次
- Deterministic Differentiable Structured Pruning for Large Language ModelsWeiyu Huang, Pengle Zhang, Xiaolu Zhang, JUN ZHOU 等ICML 2026 · 被引用 1 次
- GDP: Stabilized Neural Network Pruning via Gates with Differentiable PolarizationYi Guo, Huan Yuan, Jianchao Tan, Zhangyang Wang 等ICCV 2021 · 被引用 52 次
