KPNet: Towards Minimal Face Detector
Guanglu Song, Yu Liu, Yuhang Zang, Xiaogang Wang, Biao Leng, Qingsheng Yuan
Abstract
The small receptive field and capacity of minimal neural networks limit their performance when using them to be the backbone of detectors. In this work, we find that the appearance feature of a generic face is discriminative enough for a tiny and shallow neural network to verify from the background. And the essential barriers behind us are 1) the vague definition of the face bounding box and 2) tricky design of anchor-boxes or receptive field. Unlike most top-down methods for joint face detection and alignment, the proposed KPNet detects small facial keypoints instead of the whole face by in a bottom-up manner. It first predicts the facial landmarks from a low-resolution image via the well-designed fine-grained scale approximation and scale adaptive soft-argmax operator. Finally, the precise face bounding boxes, no matter how we define it, can be inferred from the keypoints. Without any complex head architecture or meticulous network designing, the KPNet achieves state-of-the-art accuracy on generic face detection and alignment benchmarks with only parameters, which runs at 1000fps on GPU and is easy to perform real-time on most modern front-end chips.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca975ea6-323e-444c-b020-5260c35fca98Cited by top-tier papers1
Ask how each one uses itRelated papers
- Joint Super-Resolution and Alignment of Tiny FacesYu Yin, Joseph P. Robinson, Yulun Zhang, Yun FuAAAI 2020 · 38 citations
- Attention-Driven Cropping for Very High Resolution Facial Landmark DetectionPrashanth Chandran, Derek Bradley, Markus Gross, Thabo BeelerCVPR 2020
- MogFace: Towards a Deeper Appreciation on Face DetectionYang Liu, Fei Wang, Jiankang Deng, Zhipeng Zhou et al.CVPR 2022 · 26 citations
- KeyPosS: Plug-and-Play Facial Landmark Detection through GPS-Inspired True-Range MultilaterationXu Bao, Zhi-Qi Cheng, Jun-Yan He, Wangmeng Xiang et al.ACM MM 2023 · 5 citations
- CRFace: Confidence Ranker for Model-Agnostic Face Detection RefinementNoranart Vesdapunt, Baoyuan WangCVPR 2021
