Dynamic Resolution Network
Mingjian Zhu, Kai Han, Enhua Wu, Qiulin Zhang, Ying Nie, Zhenzhong Lan, Yunhe Wang
Abstract
Deep convolutional neural networks (CNNs) are often of sophisticated design with numerous learnable parameters for the accuracy reason. To alleviate the expensive costs of deploying them on mobile devices, recent works have made huge efforts for excavating redundancy in pre-defined architectures. Nevertheless, the redundancy on the input resolution of modern CNNs has not been fully investigated, i.e., the resolution of input image is fixed. In this paper, we observe that the smallest resolution for accurately predicting the given image is different using the same neural network. To this end, we propose a novel dynamic-resolution network (DRNet) in which the input resolution is determined dynamically based on each input sample. Wherein, a resolution predictor with negligible computational costs is explored and optimized jointly with the desired network. Specifically, the predictor learns the smallest resolution that can retain and even exceed the original recognition accuracy for each image. During the inference, each input image will be resized to its predicted resolution for minimizing the overall computation burden. We then conduct extensive experiments on several benchmark networks and datasets. The results show that our DRNet can be embedded in any off-the-shelf network architecture to obtain a considerable reduction in computational complexity. For instance, DR-ResNet-50 achieves similar performance with an about 34% computation reduction, while gaining 1.4% accuracy increase with 10% computation reduction compared to the original ResNet-50 on ImageNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers15
- Adaptive Rotated Convolution for Rotated Object DetectionYifan Pu, Yiru Wang, Zhuofan Xia, Yizeng Han et al.ICCV 2023 · 154 citations
- Learning Frequency Domain Approximation for Binary Neural NetworksYixing Xu, Kai Han, Chang Xu, Yehui Tang et al.NeurIPS 2021 · 64 citations
- AdaBrowse: Adaptive Video Browser for Efficient Continuous Sign Language RecognitionLianyu Hu, Liqing Gao, Zekang Liu, Chi-Man Pun et al.ACM MM 2023 · 28 citations
- Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning ExperiencesFred Hohman, Mary Beth Kery, Donghao Ren, Dominik MoritzCHI 2024 · 27 citations
- Talaria: Interactively Optimizing Machine Learning Models for Efficient InferenceFred Hohman, Chaoqun Wang, Jinmook Lee, Jochen Görtler et al.CHI 2024 · 8 citations
Builds on12
- Glance and Focus: a Dynamic Approach to Reducing Spatial Redundancy in Image ClassificationYulin Wang, Kangchen Lv, Rui Huang, Shiji Song et al.NeurIPS 2020 · 179 citations
- Provable Filter Pruning for Efficient Neural NetworksLucas Liebenwein, Cenk Baykal, Harry Lang, Dan Feldman et al.ICLR 2020 · 161 citations
- Accelerate CNNs from Three Dimensions: A Comprehensive Pruning FrameworkWenxiao Wang, Minghao Chen, Shuai Zhao, Long Chen et al.ICML 2021 · 65 citations
- Kernel Based Progressive Distillation for Adder Neural NetworksYixing Xu, Chang Xu, Xinghao Chen, Wei Zhang et al.NeurIPS 2020 · 48 citations
- GhostNet: More Features From Cheap OperationsKai Han, Yunhe Wang, Qi Tian, Jianyuan Guo et al.CVPR 2020
Related papers
- Resolution Adaptive Networks for Efficient InferenceLe Yang, Yizeng Han, Xi Chen, Shiji Song et al.CVPR 2020
- Learning Frequency-aware Dynamic Network for Efficient Super-ResolutionWenbin Xie, Dehua Song, Chang Xu, Chunjing Xu et al.ICCV 2021 · 89 citations
- URNet: User-Resizable Residual Networks with Conditional Gating ModuleSang-Ho Lee, Simyung Chang, Nojun KwakAAAI 2020 · 11 citations
- ThumbNet: One Thumbnail Image Contains All You Need for RecognitionChen Zhao, Bernard GhanemACM MM 2020 · 14 citations
- Dynamic Dual Gating Neural NetworksFanrong Li, Gang Li, Xiangyu He, Jian ChengICCV 2021 · 40 citations
