Sparse Local Patch Transformer for Robust Face Alignment and Landmarks Inherent Relation Learning
Jiahao Xia, Weiwei Qu, Wenjian Huang, Jianguo Zhang, Xi Wang, Min Xu
摘要
Heatmap regression methods have dominated face alignment area in recent years while they ignore the inherent relation between different landmarks. In this paper, we propose a Sparse Local Patch Transformer (SLPT) for learning the inherent relation. The SLPT generates the representation of each single landmark from a local patch and aggregates them by an adaptive inherent relation based on the attention mechanism. The subpixel coordinate of each landmark is predicted independently based on the aggregated feature. Moreover, a coarse-to-fine framework is further introduced to incorporate with the SLPT, which enables the initial landmarks to gradually converge to the target facial landmarks using fine-grained features from dynamically resized local patches. Extensive experiments carried out on three popular benchmarks, including WFLW, 300W and COFW, demonstrate that the proposed method works at the state-of-the-art level with much less computational complexity by learning the inherent relation between facial landmarks. The code is available at the project website <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> https://github.com/Jiahao-UTS/SLPT-master.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Learning Motion-Robust Remote Photoplethysmography through Arbitrary Resolution VideosJianwei Li, Zitong Yu, Jingang ShiAAAI 2023 · 被引用 65 次
- FaceXFormer: A Unified Transformer for Facial AnalysisKartik Narayan, Vibashan VS, Rama Chellappa, Vishal M. PatelICCV 2025 · 被引用 16 次
- KeyPosS: Plug-and-Play Facial Landmark Detection through GPS-Inspired True-Range MultilaterationXu Bao, Zhi-Qi Cheng, Jun-Yan He, Wangmeng Xiang 等ACM MM 2023 · 被引用 5 次
- FacialFlowNet: Advancing Facial Optical Flow Estimation with a Diverse Dataset and a Decomposed ModelJianzhi Lu, Ruian He, Shili Zhou, Weimin Tan 等ACM MM 2024 · 被引用 4 次
- POPoS: Improving Efficient and Robust Facial Landmark Detection with Parallel Optimal Position SearchChong-Yang Xiang, Jun-Yan He, Zhi-Qi Cheng, Xiao Wu 等AAAI 2025 · 被引用 3 次
它引用的顶会 Paper11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Adaptive Wing Loss for Robust Face Alignment via Heatmap RegressionXinyao Wang, Liefeng Bo, Fuxin LiICCV 2019 · 被引用 293 次
- DeCaFA: Deep Convolutional Cascade for Face Alignment in the WildArnaud Dapogny, Matthieu Cord, Kevin BaillyICCV 2019 · 被引用 91 次
- Aggregation via Separation: Boosting Facial Landmark Detector With Semi-Supervised Style TranslationShengju Qian, Keqiang Sun, Wayne Wu, Chen Qian 等ICCV 2019 · 被引用 79 次
相关 Paper
- Towards Accurate Facial Landmark Detection via Cascaded TransformersHui Li, Zidong Guo, Seon-Min Rhee, Seungju Han 等CVPR 2022 · 被引用 45 次
- Attentive One-Dimensional Heatmap Regression for Facial Landmark Detection and TrackingShi Yin, Shangfei Wang, Xiaoping Chen, Enhong Chen 等ACM MM 2020 · 被引用 22 次
- FreeEnricher: Enriching Face Landmarks without Additional CostYangyu Huang, Xi Chen, Jongyoo Kim, Hao Yang 等AAAI 2023 · 被引用 3 次
- Knowing When to Quit: Selective Cascaded Regression with Patch Attention for Real-Time Face AlignmentGil Shapira, Noga Levy, Ishay Goldin, Roy Josef JevnisekACM MM 2021 · 被引用 3 次
- Dual Focus-Attention Transformer for Robust Point Cloud RegistrationKexue Fu, Mingzhi Yuan, Changwei Wang, Weiguang Pang 等CVPR 2025
