Model Rubik's Cube: Twisting Resolution, Depth and Width for TinyNets
Kai Han, Yunhe Wang, Qiulin Zhang, Wei Zhang, Chunjing Xu, Tong Zhang
摘要
To obtain excellent deep neural architectures, a series of techniques are carefully designed in EfficientNets. The giant formula for simultaneously enlarging the resolution, depth and width provides us a Rubik's cube for neural networks. So that we can find networks with high efficiency and excellent performance by twisting the three dimensions. This paper aims to explore the twisting rules for obtaining deep neural networks with minimum model sizes and computational costs. Different from the network enlarging, we observe that resolution and depth are more important than width for tiny networks. Therefore, the original method, i.e., the compound scaling in EfficientNet is no longer suitable. To this end, we summarize a tiny formula for downsizing neural architectures through a series of smaller models derived from the EfficientNet-B0 with the FLOPs constraint. Experimental results on the ImageNet benchmark illustrate that our TinyNet performs much better than the smaller version of EfficientNets using the inversed giant formula. For instance, our TinyNet-E achieves a 59.9% Top-1 accuracy with only 24M FLOPs, which is about 1.9% higher than that of the previous best MobileNetV3 with similar computational cost. Code will be available at this https URL, and this https URL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Mobile-Former: Bridging MobileNet and TransformerYinpeng Chen, Xiyang Dai, Dongdong Chen, Mengchen Liu 等CVPR 2022 · 被引用 600 次
- TopFormer: Token Pyramid Transformer for Mobile Semantic SegmentationWenqiang Zhang, Zilong Huang, Guozhong Luo, Tao Chen 等CVPR 2022 · 被引用 313 次
- Memory-efficient Patch-based Inference for Tiny Deep LearningJi Lin, Wei-Ming Chen, Han Cai, Chuang Gan 等NeurIPS 2021 · 被引用 190 次
- Accelerate CNNs from Three Dimensions: A Comprehensive Pruning FrameworkWenxiao Wang, Minghao Chen, Shuai Zhao, Long Chen 等ICML 2021 · 被引用 65 次
- Expediting Large-Scale Vision Transformer for Dense Prediction without Fine-tuningWeicong Liang, Yuhui Yuan, Henghui Ding, Xiao Luo 等NeurIPS 2022 · 被引用 52 次
它引用的顶会 Paper12
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- Efficient Residual Dense Block Search for Image Super-ResolutionDehua Song, Chang Xu, Xu Jia, Yiyi Chen 等AAAI 2020 · 被引用 145 次
- Co-Evolutionary Compression for Unpaired Image TranslationHan Shu, Yunhe Wang, Xu Jia, Kai Han 等ICCV 2019 · 被引用 93 次
- Training Binary Neural Networks through Learning with Noisy SupervisionKai Han, Yunhe Wang, Yixing Xu, Chunjing Xu 等ICML 2020 · 被引用 63 次
相关 Paper
- EtinyNet: Extremely Tiny Network for TinyMLKunran Xu, Yishi Li, Huawei Zhang, Rui Lai 等AAAI 2022 · 被引用 31 次
- Searching for Fast Model Families on Datacenter AcceleratorsSheng Li, Mingxing Tan, Ruoming Pang, Andrew Li 等CVPR 2021
- NeuralScale: Efficient Scaling of Neurons for Resource-Constrained Deep Neural NetworksEugene Lee, Chen-Yi LeeCVPR 2020
- EfficientDet: Scalable and Efficient Object DetectionMingxing Tan, Ruoming Pang, Quoc V. LeCVPR 2020
- Revisiting ResNets: Improved Training and Scaling StrategiesIrwan Bello, William Fedus, Xianzhi Du, Ekin Dogus Cubuk 等NeurIPS 2021 · 被引用 378 次
