AutoMC: Automated Model Compression Based on Domain Knowledge and Progressive Search
Chunnan Wang, Hongzhi Wang, Xiangyu Shi
Abstract
Model compression methods can reduce model complexity on the premise of maintaining acceptable performance, and thus promote the application of deep neural networks under resource constrained environments. Despite their great success, the selection of suitable compression methods and design of details of the compression scheme are difficult, requiring lots of domain knowledge as support, which is not friendly to non-expert users. To make more users easily access to the model compression scheme that best meet their needs, in this paper, we propose AutoMC, an effective and efficient automatic tool for model compression. In order to improve the search efficiency and quality, in AutoMC, we build the domain knowledge on model compression to deeply understand the characteristics and advantages of each compression method under different settings. This method can provide AutoMC with the more reasonable guidance and thus reduce useless evaluation. In addition, we present a progressive search strategy to efficiently explore pareto optimal compression scheme according to the learned prior knowledge combined with the historical evaluation information. This strategy can help AutoMC selectively and gradually explore more valuable search space, and thus reduce the search difficulty and improve the search efficiency. Extensive experimental results show that AutoMC can provide users with better compression schemes within short time compared to the existing compression methods and AutoML algorithms, which demonstrates the effectiveness and significance of our proposed algorithm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on6
- MetaPruning: Meta Learning for Automatic Neural Network Channel PruningZechun Liu, Haoyuan Mu, Xiangyu Zhang, Zichao Guo et al.ICCV 2019 · 633 citations
- Learning Filter Basis for Convolutional Neural Network CompressionYawei Li, Shuhang Gu, Luc Van Gool, Radu TimofteICCV 2019 · 106 citations
- Warm Starting CMA-ES for Hyperparameter OptimizationMasahiro Nomura, Shuhei Watanabe, Youhei Akimoto, Yoshihiko Ozaki et al.AAAI 2021 · 59 citations
- DeepLine: AutoML Tool for Pipelines Generation using Deep Reinforcement Learning and Hierarchical Actions FilteringYuval Heffetz, Roman Vainshtein, Gilad Katz, Lior RokachKDD 2020 · 3 citations
- Light Multi-Segment Activation for Model CompressionZhenhui Xu, Guolin Ke, Jia Zhang, Jiang Bian et al.AAAI 2020 · 2 citations
Related papers
- Neural Epitome Search for Architecture-Agnostic Network CompressionDaquan Zhou, Xiaojie Jin, Qibin Hou, Kaixin Wang et al.ICLR 2020 · 13 citations
- Auto Graph Encoder-Decoder for Neural Network PruningSixing Yu, Arya Mazaheri, Ali JannesariICCV 2021 · 47 citations
- AutoGAN-Distiller: Searching to Compress Generative Adversarial NetworksYonggan Fu, Wuyang Chen, Haotao Wang, Haoran Li et al.ICML 2020 · 91 citations
- Automatic Channel Pruning with Hyper-parameter Search and Dynamic MaskingBaopu Li, Yanwen Fan, Zhihong Pan, Yuchen Bian et al.ACM MM 2021 · 3 citations
- AutoCompress: An Automatic DNN Structured Pruning Framework for Ultra-High Compression RatesNing Liu, Xiaolong Ma, Zhiyuan Xu, Yanzhi Wang et al.AAAI 2020 · 204 citations
