Filter Grafting for Deep Neural Networks
Fanxu Meng, Hao Cheng, Ke Li, Zhixin Xu, Rongrong Ji, Xing Sun, Guangming Lu
Abstract
Filter is the key component in modern convolutional neural networks (CNNs). However, since CNNs are usually over-parameterized, a pre-trained network always contain some invalid (unimportant) filters. These filters have relatively small l 1 norm and contribute little to the output (Reason). While filter pruning removes these invalid filters for efficiency consideration, we tend to reactivate them to improve the representation capability of CNNs. In this paper, we introduce filter grafting (Method) to achieve this goal. The activation is processed by grafting external information (weights) into invalid filters. To better perform the grafting, we develop a novel criterion to measure the information of filters and an adaptive weighting strategy to balance the grafted information among networks. After the grafting operation, the network has fewer invalid filters compared with its initial state, enpowering the model with more representation capacity. Meanwhile, since grafting is operated reciprocally on all networks involved, we find that grafting may lose the information of valid filters when improving invalid filters. To gain a universal improvement on both valid and invalid filters, we compensate grafting with distillation (Cultivation) to overcome the drawback of grafting . Extensive experiments are performed on the classification and recognition tasks to show the superiority of our method. Code is available at https://github.com/fxmeng/filter-grafting .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 14fa21df-281a-4809-b047-ddefc91260dbCited by top-tier papers3
- Progressive Network Grafting for Few-Shot Knowledge DistillationChengchao Shen, Xinchao Wang, Youtan Yin, Jie Song et al.AAAI 2021 · 55 citations
- Weight Evolution: Improving Deep Neural Networks Training through Evolving Inferior Weight ValuesZhenquan Lin, Kailing Guo, Xiaofen Xing, Xiangmin XuACM MM 2021 · 1 citation
- CondenseNet V2: Sparse Feature Reactivation for Deep NetworksLe Yang, Haojun Jiang, Ruojin Cai, Yulin Wang et al.CVPR 2021
Builds on2
Related papers
- Provable Filter Pruning for Efficient Neural NetworksLucas Liebenwein, Cenk Baykal, Harry Lang, Dan Feldman et al.ICLR 2020 · 161 citations
- Bayesian based Re-parameterization for DNN Model PruningXiaotong Lu, Teng Xi, Baopu Li, Gang Zhang et al.ACM MM 2022 · 4 citations
- Embracing the Dark Knowledge: Domain Generalization Using Regularized Knowledge DistillationYufei Wang, Haoliang Li, Lap-Pui Chau, Alex C. KotACM MM 2021 · 46 citations
- Neural Network Pruning With Residual-Connections and Limited-DataJian-Hao Luo, Jianxin WuCVPR 2020
- Linearly Replaceable Filters for Deep Network Channel PruningDonggyu Joo, Eojindl Yi, Sunghyun Baek, Junmo KimAAAI 2021 · 39 citations
