Learning Robust Facial Landmark Detection via Hierarchical Structured Ensemble
Xu Zou, Sheng Zhong, Luxin Yan, Xiangyun Zhao, Jiahuan Zhou, Ying Wu
Abstract
Heatmap regression-based models have significantly advanced the progress of facial landmark detection. However, the lack of structural constraints always generates inaccurate heatmaps resulting in poor landmark detection performance. While hierarchical structure modeling methods have been proposed to tackle this issue, they all heavily rely on manually designed tree structures. The designed hierarchical structure is likely to be completely corrupted due to the missing or inaccurate prediction of landmarks. To the best of our knowledge, in the context of deep learning, no work before has investigated how to automatically model proper structures for facial landmarks, by discovering their inherent relations. In this paper, we propose a novel Hierarchical Structured Landmark Ensemble (HSLE) model for learning robust facial landmark detection, by using it as the structural constraints. Different from existing approaches of manually designing structures, our proposed HSLE model is constructed automatically via discovering the most robust patterns so HSLE has the ability to robustly depict both local and holistic landmark structures simultaneously. Our proposed HSLE can be readily plugged into any existing facial landmark detection baselines for further performance improvement. Extensive experimental results demonstrate our approach significantly outperforms the baseline by a large margin to achieve a state-of-the-art performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6ab8140f-408a-4c4a-bc44-2ad263c80361Cited by top-tier papers12
- General Facial Representation Learning in a Visual-Linguistic MannerYinglin Zheng, Hao Yang, Ting Zhang, Jianmin Bao et al.CVPR 2022 · 161 citations
- Sparse Local Patch Transformer for Robust Face Alignment and Landmarks Inherent Relation LearningJiahao Xia, Weiwei Qu, Wenjian Huang, Jianguo Zhang et al.CVPR 2022 · 50 citations
- Towards Accurate Facial Landmark Detection via Cascaded TransformersHui Li, Zidong Guo, Seon-Min Rhee, Seungju Han et al.CVPR 2022 · 45 citations
- Occlusion-robust Face Alignment using A Viewpoint-invariant Hierarchical Network ArchitectureCongcong Zhu, Xintong Wan, Shaorong Xie, Xiaoqiang Li et al.CVPR 2022 · 16 citations
- KeyPosS: Plug-and-Play Facial Landmark Detection through GPS-Inspired True-Range MultilaterationXu Bao, Zhi-Qi Cheng, Jun-Yan He, Wangmeng Xiang et al.ACM MM 2023 · 5 citations
Related papers
- Heatmap Regression without Soft-Argmax for Facial Landmark DetectionChiao-An Yang, Raymond A. YehICCV 2025 · 3 citations
- Learning to Detect 3D Facial Landmarks via Heatmap Regression with Graph Convolutional NetworkYuan Wang, Min Cao, Zhenfeng Fan, Silong PengAAAI 2022 · 30 citations
- PossLoss: A Reliable and Sensitive Facial Landmark Detection Loss FunctionQikui ZhuICCV 2025 · 1 citation
- Attentive One-Dimensional Heatmap Regression for Facial Landmark Detection and TrackingShi Yin, Shangfei Wang, Xiaoping Chen, Enhong Chen et al.ACM MM 2020 · 22 citations
- STAR Loss: Reducing Semantic Ambiguity in Facial Landmark DetectionZhenglin Zhou, Huaxia Li, Hong Liu, Nanyang Wang et al.CVPR 2023
