DeCaFA: Deep Convolutional Cascade for Face Alignment in the Wild
Arnaud Dapogny, Matthieu Cord, Kevin Bailly
Abstract
Face Alignment is an active computer vision domain, that consists in localizing a number of facial landmarks that vary across datasets. State-of-the-art face alignment methods either consist in end-to-end regression, or in refining the shape in a cascaded manner, starting from an initial guess. In this paper, we introduce an end-to-end deep convolutional cascade (DeCaFA) architecture for face alignment. Face Alignment is an active computer vision domain, that consists in localizing a number of facial landmarks that vary across datasets. State-of-the-art face alignment methods either consist in end-to-end regression, or in refining the shape in a cascaded manner, starting from an initial guess. In this paper, we introduce DeCaFA, an end-to-end deep convolutional cascade architecture for face alignment. DeCaFA uses fully-convolutional stages to keep full spatial resolution throughout the cascade. Between each cascade stage, DeCaFA uses multiple chained transfer layers with spatial softmax to produce landmark-wise attention maps for each of several landmark alignment tasks. Weighted intermediate supervision, as well as efficient feature fusion between the stages allow to learn to progressively refine the attention maps in an end-to-end manner. We show experimentally that DeCaFA significantly outperforms existing approaches on 300W, CelebA and WFLW databases. In addition, we show that DeCaFA can learn fine alignment with reasonable accuracy from very few images using coarsely annotated data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f0e1c1ca-e3e3-49fb-af2c-62fee40669e6Cited by top-tier papers12
- General Facial Representation Learning in a Visual-Linguistic MannerYinglin Zheng, Hao Yang, Ting Zhang, Jianmin Bao et al.CVPR 2022 · 161 citations
- ADNet: Leveraging Error-Bias Towards Normal Direction in Face AlignmentYangyu Huang, Hao Yang, Chong Li, Jongyoo Kim et al.ICCV 2021 · 63 citations
- Sparse Local Patch Transformer for Robust Face Alignment and Landmarks Inherent Relation LearningJiahao Xia, Weiwei Qu, Wenjian Huang, Jianguo Zhang et al.CVPR 2022 · 50 citations
- Towards Accurate Facial Landmark Detection via Cascaded TransformersHui Li, Zidong Guo, Seon-Min Rhee, Seungju Han et al.CVPR 2022 · 45 citations
- FaceXFormer: A Unified Transformer for Facial AnalysisKartik Narayan, Vibashan VS, Rama Chellappa, Vishal M. PatelICCV 2025 · 16 citations
Related papers
- Heatmap Regression without Soft-Argmax for Facial Landmark DetectionChiao-An Yang, Raymond A. YehICCV 2025 · 3 citations
- FAN-Face: a Simple Orthogonal Improvement to Deep Face RecognitionJing Yang, Adrian Bulat, Georgios TzimiropoulosAAAI 2020 · 28 citations
- FreeEnricher: Enriching Face Landmarks without Additional CostYangyu Huang, Xi Chen, Jongyoo Kim, Hao Yang et al.AAAI 2023 · 3 citations
- Learning to Detect 3D Facial Landmarks via Heatmap Regression with Graph Convolutional NetworkYuan Wang, Min Cao, Zhenfeng Fan, Silong PengAAAI 2022 · 30 citations
- RetinaFace: Single-Shot Multi-Level Face Localisation in the WildJiankang Deng, Jia Guo, Evangelos Ververas, Irene Kotsia et al.CVPR 2020
