DamoFD: Digging into Backbone Design on Face Detection
Yang Liu, Jiankang Deng, Fei Wang, Lei Shang, Xuansong Xie, Baigui Sun
Abstract
Face detection (FD) has achieved remarkable success over the past few years, yet, these leaps often arrive when consuming enormous computation costs. Moreover, when considering a realistic situation, i.e., building a lightweight face detector under a computation-scarce scenario, such heavy computation cost limits the application of the face detector. To remedy this, several pioneering works design tiny face detectors through off-the-shelf neural architecture search (NAS) technologies, which are usually applied to the classification task. Thus, the searched architectures are sub-optimal for the face detection task since some design criteria between detection and classification task are different. As a representative, the face detection backbone design needs to guarantee the stage-level detection ability while it is not required for the classification backbone. Furthermore, the detection backbone consumes a vast body of inference budgets in the whole detection framework. Considering the intrinsic design requirement and the virtual importance role of the face detection backbone, we thus ask a critical question: How to employ NAS to search FD-friendly backbone architecture? To cope with this question, we propose a distribution-dependent stage-aware ranking score (DDSAR-Score) to explicitly characterize the stage-level expressivity and identify the individual importance of each stage, thus satisfying the aforementioned design criterion of the FD backbone. Based on our proposed DDSAR-Score, we conduct comprehensive experiments on the challenging Wider Face benchmark dataset and achieve dominant performance across a wide range of compute regimes. In particular, compared to the tiniest face detector SCRFD-0.5GF, our method is +2.5 % better in Average Precision (AP) score when using the same amount of FLOPs. The
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Sample and Computation Redistribution for Efficient Face DetectionJia Guo, Jiankang Deng, Alexandros Lattas, Stefanos ZafeiriouICLR 2022 · 173 citations
- Zen-NAS: A Zero-Shot NAS for High-Performance Image RecognitionMing Lin, Pichao Wang, Zhenhong Sun, Hesen Chen et al.ICCV 2021 · 164 citations
- On the Number of Linear Regions of Convolutional Neural NetworksHuan Xiong, Lei Huang, Mengyang Yu, Li Liu et al.ICML 2020 · 80 citations
- MogFace: Towards a Deeper Appreciation on Face DetectionYang Liu, Fei Wang, Jiankang Deng, Zhipeng Zhou et al.CVPR 2022 · 26 citations
- ASFD: Automatic and Scalable Face DetectorJian Li, Bin Zhang, Yabiao Wang, Ying Tai et al.ACM MM 2021 · 20 citations
Related papers
- BFBox: Searching Face-Appropriate Backbone and Feature Pyramid Network for Face DetectorYang Liu, Xu TangCVPR 2020
- MAE-DET: Revisiting Maximum Entropy Principle in Zero-Shot NAS for Efficient Object DetectionZhenhong Sun, Ming Lin, Xiuyu Sun, Zhiyu Tan et al.ICML 2022 · 40 citations
- Computation Reallocation for Object DetectionFeng Liang, Chen Lin, Ronghao Guo, Ming Sun et al.ICLR 2020 · 36 citations
- SM-NAS: Structural-to-Modular Neural Architecture Search for Object DetectionLewei Yao, Hang Xu, Wei Zhang, Xiaodan Liang et al.AAAI 2020 · 83 citations
- Entropy-Driven Mixed-Precision Quantization for Deep Network DesignZhenhong Sun, Ce Ge, Junyan Wang, Ming Lin et al.NeurIPS 2022 · 41 citations
