Learning Spatial-Semantic Relationship for Facial Attribute Recognition With Limited Labeled Data
Ying Shu, Yan Yan, Si Chen, Jing-Hao Xue, Chunhua Shen, Hanzi Wang
Abstract
Recent advances in deep learning have demonstrated excellent results for Facial Attribute Recognition (FAR), typically trained with large-scale labeled data. However, in many real-world FAR applications, only limited labeled data are available, leading to remarkable deterioration in performance for most existing deep learning-based FAR methods. To address this problem, here we propose a method termed Spatial-Semantic Patch Learning (SSPL). The training of SSPL involves two stages. First, three auxiliary tasks, consisting of a Patch Rotation Task (PRT), a Patch Segmentation Task (PST), and a Patch Classification Task (PC-T), are jointly developed to learn the spatial-semantic relationship from large-scale unlabeled facial data. We thus obtain a powerful pre-trained model. In particular, PRT exploits the spatial information of facial images in a selfsupervised learning manner. PST and PCT respectively capture the pixel-level and image-level semantic information of facial images based on a facial parsing model. Second, the spatial-semantic knowledge learned from auxiliary tasks is transferred to the FAR task. By doing so, it enables that only a limited number of labeled data are required to fine-tune the pre-trained model. We achieve superior performance compared with state-of-the-art methods, as substantiated by extensive experiments and studies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c70c8fd0-9dae-4e17-8b49-59994dca44c3Cited by top-tier papers7
- General Facial Representation Learning in a Visual-Linguistic MannerYinglin Zheng, Hao Yang, Ting Zhang, Jianmin Bao et al.CVPR 2022 · 161 citations
- Hierarchical Visual Primitive Experts for Compositional Zero-Shot LearningHanjae Kim, Jiyoung Lee, Seongheon Park, Kwanghoon SohnICCV 2023 · 27 citations
- Enhancing Face Recognition with Self-Supervised 3D ReconstructionMingjie He, Jie Zhang, Shiguang Shan, Xilin ChenCVPR 2022 · 26 citations
- MogFace: Towards a Deeper Appreciation on Face DetectionYang Liu, Fei Wang, Jiankang Deng, Zhipeng Zhou et al.CVPR 2022 · 26 citations
- FaceXFormer: A Unified Transformer for Facial AnalysisKartik Narayan, Vibashan VS, Rama Chellappa, Vishal M. PatelICCV 2025 · 16 citations
Builds on2
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
Related papers
- LAFS: Landmark-Based Facial Self-Supervised Learning for Face RecognitionZhonglin Sun, Chen Feng, Ioannis Patras, Georgios TzimiropoulosCVPR 2024 · 17 citations
- Self-Supervised Regional and Temporal Auxiliary Tasks for Facial Action Unit RecognitionJingwei Yan, Jingjing Wang, Qiang Li, Chunmao Wang et al.ACM MM 2021 · 9 citations
- Exploiting Self-Supervised and Semi-Supervised Learning for Facial Landmark Tracking with Unlabeled DataShi Yin, Shangfei Wang, Xiaoping Chen, Enhong ChenACM MM 2020 · 7 citations
- Knowledge-Driven Self-Supervised Representation Learning for Facial Action Unit RecognitionYanan Chang, Shangfei WangCVPR 2022 · 38 citations
- PrefAce: Face-Centric Pretraining with Self-Structure Aware DistillationSiyuan Hu, Zheng Wang, Peng Hu, Xi Peng et al.AAAI 2024 · 2 citations
