Exploiting Semantic Embedding and Visual Feature for Facial Action Unit Detection
Huiyuan Yang, Lijun Yin, Yi Zhou, Jiuxiang Gu
Abstract
Recent study on detecting facial action units (AU) has utilized auxiliary information (i.e., facial landmarks, relationship among AUs and expressions, web facial images, etc.), in order to improve the AU detection performance. As of now, no semantic information of AUs has yet been explored for such a task. As a matter of fact, AU semantic descriptions provide much more information than the binary AU labels alone, thus we propose to exploit the Semantic Embedding and Visual feature (SEV-Net) for AU detection. More specifically, AU semantic embeddings are obtained through both Intra-AU and Inter-AU attention modules, where the Intra-AU attention module captures the relation among words within each sentence that describes individual AU, and the Inter-AU attention module focuses on the relation among those sentences. The learned AU semantic embeddings are then used as guidance for the generation of attention maps through a cross-modality attention network. The generated cross-modality attention maps are further used as weights for the aggregated feature. Our proposed method is unique in that the semantic features are exploited as the first of this kind. The approach has been evaluated on three public AU-coded facial expression databases, and has achieved a superior performance than the state-of-the-art peer methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a43fbdc4-0dfd-4a34-9461-a12db50c044bCited by top-tier papers10
- Causal Intervention for Subject-Deconfounded Facial Action Unit RecognitionYingjie Chen, Diqi Chen, Tao Wang, Yizhou Wang et al.AAAI 2022 · 36 citations
- Context-Aware Feature and Label Fusion for Facial Action Unit Intensity Estimation With Partially Labeled DataYong Zhang, Haiyong Jiang, Baoyuan Wu, Yanbo Fan et al.ICCV 2019 · 32 citations
- Weakly-Supervised Text-driven Contrastive Learning for Facial Behavior UnderstandingXiang Zhang, Taoyue Wang, Xiaotian Li, Huiyuan Yang et al.ICCV 2023 · 26 citations
- Knowledge-Spreader: Learning Semi-Supervised Facial Action Dynamics by Consistifying Knowledge GranularityXiaotian Li, Xiang Zhang, Taoyue Wang, Lijun YinICCV 2023 · 18 citations
- Towards End-to-End Explainable Facial Action Unit Recognition via Vision-Language Joint LearningXuri Ge, Junchen Fu, Fuhai Chen, Shan An et al.ACM MM 2024 · 12 citations
Builds on2
Related papers
- Facial Action Unit Intensity Estimation via Semantic Correspondence Learning with Dynamic Graph ConvolutionYingruo Fan, Jacqueline C. K. Lam, Victor On Kwok LiAAAI 2020 · 58 citations
- Integrating Semantic and Temporal Relationships in Facial Action Unit DetectionZhihua Li, Xiang Deng, Xiaotian Li, Lijun YinACM MM 2021 · 11 citations
- Facial Action Unit Detection With TransformersGeethu Miriam Jacob, Björn StengerCVPR 2021
- Region of Interest Based Graph Convolution: A Heatmap Regression Approach for Action Unit DetectionZheng Zhang, Taoyue Wang, Lijun YinACM MM 2020 · 21 citations
- Knowledge Augmented Deep Neural Networks for Joint Facial Expression and Action Unit RecognitionZijun Cui, Tengfei Song, Yuru Wang, Qiang JiNeurIPS 2020 · 70 citations
