Attention-Driven Cropping for Very High Resolution Facial Landmark Detection
Prashanth Chandran, Derek Bradley, Markus Gross, Thabo Beeler
摘要
Facial landmark detection is a fundamental task for many consumer and high-end applications and is almost entirely solved by machine learning methods today. Existing datasets used to train such algorithms are primarily made up of only low resolution images, and current algorithms are limited to inputs of comparable quality and resolution as the training dataset. On the other hand, high resolution imagery is becoming increasingly more common as consumer cameras improve in quality every year. Therefore, there is need for algorithms that can leverage the rich information available in high resolution imagery. Naïvely attempting to reuse existing network architectures on high resolution imagery is prohibitive due to memory bottlenecks on GPUs. The only current solution is to downsample the images, sacrificing resolution and quality. Building on top of recent progress in attention-based networks, we present a novel, fully convolutional regional architecture that is specially designed for predicting landmarks on very high resolution facial images without downsampling. We demonstrate the flexibility of our architecture by training the proposed model with images of resolutions ranging from 256 x 256 to 4K. In addition to being the first method for facial landmark detection on high resolution images, our approach achieves superior performance over traditional (holistic) state-of-the-art architectures across ALL resolutions, leading to a general-purpose, extremely flexible, high quality landmark detector.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Towards Accurate Facial Landmark Detection via Cascaded TransformersHui Li, Zidong Guo, Seon-Min Rhee, Seungju Han 等CVPR 2022 · 被引用 45 次
- Improving Robustness of Facial Landmark Detection by Defending against Adversarial AttacksCongcong Zhu, Xiaoqiang Li, Jide Li, Songmin DaiICCV 2021 · 被引用 34 次
- Localization with Sampling-ArgmaxJiefeng Li, Tong Chen, Ruiqi Shi, Yujing Lou 等NeurIPS 2021 · 被引用 25 次
- KeyPosS: Plug-and-Play Facial Landmark Detection through GPS-Inspired True-Range MultilaterationXu Bao, Zhi-Qi Cheng, Jun-Yan He, Wangmeng Xiang 等ACM MM 2023 · 被引用 5 次
- POPoS: Improving Efficient and Robust Facial Landmark Detection with Parallel Optimal Position SearchChong-Yang Xiang, Jun-Yan He, Zhi-Qi Cheng, Xiao Wu 等AAAI 2025 · 被引用 3 次
它引用的顶会 Paper2
相关 Paper
- Attentive One-Dimensional Heatmap Regression for Facial Landmark Detection and TrackingShi Yin, Shangfei Wang, Xiaoping Chen, Enhong Chen 等ACM MM 2020 · 被引用 22 次
- Joint Super-Resolution and Alignment of Tiny FacesYu Yin, Joseph P. Robinson, Yulun Zhang, Yun FuAAAI 2020 · 被引用 38 次
- Heatmap Regression without Soft-Argmax for Facial Landmark DetectionChiao-An Yang, Raymond A. YehICCV 2025 · 被引用 3 次
- Learning to Detect 3D Facial Landmarks via Heatmap Regression with Graph Convolutional NetworkYuan Wang, Min Cao, Zhenfeng Fan, Silong PengAAAI 2022 · 被引用 30 次
- Towards High-Resolution Salient Object DetectionYi Zeng, Pingping Zhang, Zhe Lin, Jianming Zhang 等ICCV 2019 · 被引用 232 次
