MAE-DET: Revisiting Maximum Entropy Principle in Zero-Shot NAS for Efficient Object Detection
Zhenhong Sun, Ming Lin, Xiuyu Sun, Zhiyu Tan, Hao Li, Rong Jin
Abstract
In object detection, the detection backbone consumes more than half of the overall inference cost. Recent researches attempt to reduce this cost by optimizing the backbone architecture with the help of Neural Architecture Search (NAS). However, existing NAS methods for object detection require hundreds to thousands of GPU hours of searching, making them impractical in fast-paced research and development. In this work, we propose a novel zero-shot NAS method to address this issue. The proposed method, named MAE-DET, automatically designs efficient detection backbones via the Maximum Entropy Principle without training network parameters, reducing the architecture design cost to nearly zero yet delivering the state-of-the-art (SOTA) performance. Under the hood, MAE-DET maximizes the differential entropy of detection backbones, leading to a better feature extractor for object detection under the same computational budgets. After merely one GPU day of fully automatic design, MAE-DET innovates SOTA detection backbones on multiple detection benchmark datasets with little human intervention. Comparing to ResNet-50 backbone, MAE-DET is better in mAP when using the same amount of FLOPs/parameters, and is times faster on NVIDIA V100 at the same mAP. Code and pre-trained models are available at https://github.com/alibaba/lightweight-neuralarchitecture-search.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- EMQ: Evolving Training-free Proxies for Automated Mixed Precision QuantizationPeijie Dong, Lujun Li, Zimian Wei, Xin Niu et al.ICCV 2023 · 51 citations
- Entropy-Driven Mixed-Precision Quantization for Deep Network DesignZhenhong Sun, Ce Ge, Junyan Wang, Ming Lin et al.NeurIPS 2022 · 41 citations
- TransFace: Calibrating Transformer Training for Face Recognition from a Data-Centric PerspectiveJun Dan, Yang Liu, Haoyu Xie, Jiankang Deng et al.ICCV 2023 · 36 citations
- CE-NAS: An End-to-End Carbon-Efficient Neural Architecture Search FrameworkYiyang Zhao, Yunzhuo Liu, Bo Jiang, Tian GuoNeurIPS 2024 · 10 citations
- On skip connections and normalisation layers in deep optimisationLachlan E. MacDonald, Jack Valmadre, Hemanth Saratchandran, Simon LuceyNeurIPS 2023 · 8 citations
Builds on18
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- Generalized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object DetectionXiang Li, Wenhai Wang, Lijun Wu, Shuo Chen et al.NeurIPS 2020 · 2,118 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 1,188 citations
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
Related papers
- NAS-FCOS: Fast Neural Architecture Search for Object DetectionNing Wang, Yang Gao, Hao Chen, Peng Wang et al.CVPR 2020
- Hit-Detector: Hierarchical Trinity Architecture Search for Object DetectionJianyuan Guo, Kai Han, Yunhe Wang, Chao Zhang et al.CVPR 2020
- Fast Neural Network Adaptation via Parameter Remapping and Architecture SearchJiemin Fang, Yuzhu Sun, Kangjian Peng, Qian Zhang et al.ICLR 2020 · 36 citations
- SM-NAS: Structural-to-Modular Neural Architecture Search for Object DetectionLewei Yao, Hang Xu, Wei Zhang, Xiaodan Liang et al.AAAI 2020 · 83 citations
- SP-NAS: Serial-to-Parallel Backbone Search for Object DetectionChenhan Jiang, Hang Xu, Wei Zhang, Xiaodan Liang et al.CVPR 2020
