Context-Aware Feature and Label Fusion for Facial Action Unit Intensity Estimation With Partially Labeled Data
Yong Zhang, Haiyong Jiang, Baoyuan Wu, Yanbo Fan, Qiang Ji
Abstract
Facial action unit (AU) intensity estimation is a fundamental task for facial behaviour analysis. Most previous methods use a whole face image as input for intensity prediction. Considering that AUs are defined according to their corresponding local appearance, a few patch-based methods utilize image features of local patches. However, fusion of local features is always performed via straightforward feature concatenation or summation. Besides, these methods require fully annotated databases for model learning, which is expensive to acquire. In this paper, we propose a novel weakly supervised patch-based deep model on basis of two types of attention mechanisms for joint intensity estimation of multiple AUs. The model consists of a feature fusion module and a label fusion module. And we augment attention mechanisms of these two modules with a learnable task-related context, as one patch may play different roles in analyzing different AUs and each AU has its own temporal evolution rule. The context-aware feature fusion module is used to capture spatial relationships among local patches while the context-aware label fusion module is used to capture the temporal dynamics of AUs. The latter enables the model to be trained on a partially annotated database. Experimental evaluations on two benchmark expression databases demonstrate the superior performance of the proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6f3be690-97b3-43f6-a73f-5dd4ef4e663dCited by top-tier papers6
- Knowledge Augmented Deep Neural Networks for Joint Facial Expression and Action Unit RecognitionZijun Cui, Tengfei Song, Yuru Wang, Qiang JiNeurIPS 2020 · 70 citations
- Self-Supervised Regional and Temporal Auxiliary Tasks for Facial Action Unit RecognitionJingwei Yan, Jingjing Wang, Qiang Li, Chunmao Wang et al.ACM MM 2021 · 9 citations
- Unsupervised Learning Facial Parameter Regressor for Action Unit Intensity Estimation via Differentiable RendererXinhui Song, Tianyang Shi, Zunlei Feng, Mingli Song et al.ACM MM 2020 · 6 citations
- Trend-Aware Supervision: On Learning Invariance for Semi-supervised Facial Action Unit Intensity EstimationYingjie Chen, Jiarui Zhang, Tao Wang, Yun LiangAAAI 2024 · 1 citation
- Affective Processes: Stochastic Modelling of Temporal Context for Emotion and Facial Expression RecognitionEnrique Sanchez, Mani Kumar Tellamekala, Michel F. Valstar, Georgios TzimiropoulosCVPR 2021
Builds on22
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
Related papers
- Facial Action Unit Intensity Estimation via Semantic Correspondence Learning with Dynamic Graph ConvolutionYingruo Fan, Jacqueline C. K. Lam, Victor On Kwok LiAAAI 2020 · 58 citations
- Dynamic Probabilistic Graph Convolution for Facial Action Unit Intensity EstimationTengfei Song, Zijun Cui, Yuru Wang, Wenming Zheng et al.CVPR 2021
- CaFGraph: Context-aware Facial Multi-graph Representation for Facial Action Unit RecognitionYingjie Chen, Diqi Chen, Yizhou Wang, Tao Wang et al.ACM MM 2021 · 10 citations
- Integrating Semantic and Temporal Relationships in Facial Action Unit DetectionZhihua Li, Xiang Deng, Xiaotian Li, Lijun YinACM MM 2021 · 11 citations
- Exploiting Semantic Embedding and Visual Feature for Facial Action Unit DetectionHuiyuan Yang, Lijun Yin, Yi Zhou, Jiuxiang GuCVPR 2021
