TransFER: Learning Relation-aware Facial Expression Representations with Transformers
Fanglei Xue, Qiangchang Wang, Guodong Guo
摘要
Facial expression recognition (FER) has received increasing interest in computer vision. We propose the Trans-FER model which can learn rich relation-aware local representations. It mainly consists of three components: Multi-Attention Dropping (MAD), ViT-FER, and Multi-head Self-Attention Dropping (MSAD). First, local patches play an important role in distinguishing various expressions, however, few existing works can locate discriminative and diverse local patches. This can cause serious problems when some patches are invisible due to pose variations or viewpoint changes. To address this issue, the MAD is proposed to randomly drop an attention map. Consequently, models are pushed to explore diverse local patches adaptively. Second, to build rich relations between different local patches, the Vision Transformers (ViT) are used in FER, called ViT-FER. Since the global scope is used to reinforce each local patch, a better representation is obtained to boost the FER performance. Thirdly, the multi-head self-attention allows ViT to jointly attend to features from different information subspaces at different positions. Given no explicit guidance, however, multiple self-attentions may extract similar relations. To address this, the MSAD is proposed to randomly drop one self-attention module. As a result, models are forced to learn rich relations among diverse local patches. Our proposed TransFER model outperforms the state-of-the-art methods on several FER benchmarks, showing its effectiveness and usefulness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Face2Exp: Combating Data Biases for Facial Expression RecognitionDan Zeng, Zhiyuan Lin, Xiao Yan, Yuting Liu 等CVPR 2022 · 被引用 125 次
- Towards Semi-Supervised Deep Facial Expression Recognition with An Adaptive Confidence MarginHangyu Li, Nannan Wang, Xi Yang, Xiaoyu Wang 等CVPR 2022 · 被引用 97 次
- Intensity-Aware Loss for Dynamic Facial Expression Recognition in the WildHanting Li, Hongjing Niu, Zhaoqing Zhu, Feng ZhaoAAAI 2023 · 被引用 94 次
- LA-Net: Landmark-Aware Learning for Reliable Facial Expression Recognition under Label NoiseZhiyu Wu, Jinshi CuiICCV 2023 · 被引用 47 次
- Leave No Stone Unturned: Mine Extra Knowledge for Imbalanced Facial Expression RecognitionYuhang Zhang, Yaqi Li, Lixiong Qin, Xuannan Liu 等NeurIPS 2023 · 被引用 47 次
它引用的顶会 Paper5
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Label Distribution Learning on Auxiliary Label Space Graphs for Facial Expression RecognitionShikai Chen, Jianfeng Wang, Yuedong Chen, Zhongchao Shi 等CVPR 2020
- Transformer Interpretability Beyond Attention VisualizationHila Chefer, Shir Gur, Lior WolfCVPR 2021
- Suppressing Uncertainties for Large-Scale Facial Expression RecognitionKai Wang, Xiaojiang Peng, Jianfei Yang, Shijian Lu 等CVPR 2020
相关 Paper
- MAE-DFER: Efficient Masked Autoencoder for Self-supervised Dynamic Facial Expression RecognitionLicai Sun, Zheng Lian, Bin Liu, Jianhua TaoACM MM 2023 · 被引用 85 次
- TransFG: A Transformer Architecture for Fine-Grained RecognitionJu He, Jieneng Chen, Shuai Liu, Adam Kortylewski 等AAAI 2022 · 被引用 529 次
- Latent-OFER: Detect, Mask, and Reconstruct with Latent Vectors for Occluded Facial Expression RecognitionIsack Lee, Eungi Lee, Seok Bong YooICCV 2023 · 被引用 41 次
- EViT: Expediting Vision Transformers via Token ReorganizationsYouwei Liang, Chongjian Ge, Zhan Tong, Yibing Song 等ICLR 2022 · 被引用 137 次
- HKAFER: Achieve Visual Parameter-Efficient Fine-Tuning via Heterogeneous Kronecker Adaptation for Facial Expression RecognitionYu Gao, Haoyu Ji, Zhiyong Wang, Wenze Huang 等AAAI 2026
