Self-Supervised Regional and Temporal Auxiliary Tasks for Facial Action Unit Recognition
Jingwei Yan, Jingjing Wang, Qiang Li, Chunmao Wang, Shiliang Pu
Abstract
Automatic facial action unit (AU) recognition is a challenging task due to the scarcity of manual annotations. To alleviate this problem, a large amount of efforts has been dedicated to exploiting various methods which leverage numerous unlabeled data. However, many aspects with regard to some unique properties of AUs, such as the regional and relational characteristics, are not sufficiently explored in previous works. Motivated by this, we take the AU properties into consideration and propose two auxiliary AU related tasks to bridge the gap between limited annotations and the model performance in a self-supervised manner via the unlabeled data. Specifically, to enhance the discrimination of regional features with AU relation embedding, we design a task of RoI inpainting to recover the randomly cropped AU patches. Meanwhile, a single image based optical flow estimation task is proposed to leverage the dynamic change of facial muscles and encode the motion information into the global feature representation. Based on these two self-supervised auxiliary tasks, local features, mutual relation and motion cues of AUs are better captured in the backbone network with the proposed regional and temporal based auxiliary task learning (RTATL) framework. Extensive experiments on BP4D and DISFA demonstrate the superiority of our method and new state-of-the-art performances are achieved.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3a5fe949-7b5d-410f-849f-3ff65036c7c3Cited by top-tier papers1
Ask how each one uses itBuilds on3
- S4L: Self-Supervised Semi-Supervised LearningLucas Beyer, Xiaohua Zhai, Avital Oliver, Alexander KolesnikovICCV 2019 · 854 citations
- Uncertain Graph Neural Networks for Facial Action Unit DetectionTengfei Song, Lisha Chen, Wenming Zheng, Qiang JiAAAI 2021 · 86 citations
- Context-Aware Feature and Label Fusion for Facial Action Unit Intensity Estimation With Partially Labeled DataYong Zhang, Haiyong Jiang, Baoyuan Wu, Yanbo Fan et al.ICCV 2019 · 32 citations
Related papers
- Integrating Semantic and Temporal Relationships in Facial Action Unit DetectionZhihua Li, Xiang Deng, Xiaotian Li, Lijun YinACM MM 2021 · 11 citations
- PIAP-DF: Pixel-Interested and Anti Person-Specific Facial Action Unit Detection Net with Discrete Feedback LearningYang Tang, Wangding Zeng, Dafei Zhao, Honggang ZhangICCV 2021 · 39 citations
- Region of Interest Based Graph Convolution: A Heatmap Regression Approach for Action Unit DetectionZheng Zhang, Taoyue Wang, Lijun YinACM MM 2020 · 21 citations
- Knowledge-Driven Self-Supervised Representation Learning for Facial Action Unit RecognitionYanan Chang, Shangfei WangCVPR 2022 · 38 citations
- Multi-Scale Dynamic and Hierarchical Relationship Modeling for Facial Action Units RecognitionZihan Wang, Siyang Song, Cheng Luo, Songhe Deng et al.CVPR 2024
