Latent-OFER: Detect, Mask, and Reconstruct with Latent Vectors for Occluded Facial Expression Recognition
Isack Lee, Eungi Lee, Seok Bong Yoo
Abstract
Most research on facial expression recognition (FER) is conducted in highly controlled environments, but its performance is often unacceptable when applied to real-world situations. This is because when unexpected objects occlude the face, the FER network faces difficulties extracting facial features and accurately predicting facial expressions. Therefore, occluded FER (OFER) is a challenging problem. Previous studies on occlusion-aware FER have typically required fully annotated facial images for training. However, collecting facial images with various occlusions and expression annotations is time-consuming and expensive. Latent-OFER, the proposed method, can detect occlusions, restore occluded parts of the face as if they were unoccluded, and recognize them, improving FER accuracy. This approach involves three steps: First, the vision transformer (ViT)based occlusion patch detector masks the occluded position by training only latent vectors from the unoccluded patches using the support vector data description algorithm. Second, the hybrid reconstruction network generates the masking position as a complete image using the ViT and convolutional neural network (CNN). Last, the expression-relevant latent vector extractor retrieves and uses expression-related information from all latent vectors by applying a CNN-based class activation map. This mechanism has a significant advantage in preventing performance degradation from occlusion by unseen objects. The experimental results on several databases demonstrate the superiority of the proposed method over state-of-the-art methods. This code is available at https://github.com/leeisack/Latent-OFER.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bbafb887-4533-451e-b726-fefc19bde2c4Cited by top-tier papers3
- BAH Dataset for Ambivalence/Hesitancy Recognition in Videos for Digital Behavioural ChangeManuela González-González, Soufiane Belharbi, Muhammad Osama Zeeshan, Masoumeh Sharafi et al.ICLR 2026 · 18 citations
- Rethinking Occlusion in FER: A Semantic-Aware Perspective and Go BeyondHuiyu Zhai, Xingxing Yang, Yalan Ye, Chenyang Li et al.ACM MM 2025 · 5 citations
- D^3FER: Dual Channel and Dual Branch Network for Robust Facial Expression Recognition under Dual ChallengesHui Tang, Yifan He, Zhong JinCVPR 2026
Builds on13
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Intriguing Properties of Vision TransformersMuzammal Naseer, Kanchana Ranasinghe, Salman Khan, Munawar Hayat et al.NeurIPS 2021 · 863 citations
- Deep Semi-Supervised Anomaly DetectionLukas Ruff, Robert A. Vandermeulen, Nico Görnitz, Alexander Binder et al.ICLR 2020 · 678 citations
- Coherent Semantic Attention for Image InpaintingHongyu Liu, Bin Jiang, Yi Xiao, Chao YangICCV 2019 · 395 citations
- MAT: Mask-Aware Transformer for Large Hole Image InpaintingWenbo Li, Zhe Lin, Kun Zhou, Lu Qi et al.CVPR 2022 · 382 citations
Related papers
- Hypergraph-Guided Disentangled Spectrum Transformer Networks for Near-Infrared Facial Expression RecognitionBingjun Luo, Haowen Wang, Jinpeng Wang, Junjie Zhu et al.AAAI 2024 · 5 citations
- Former-DFER: Dynamic Facial Expression Recognition TransformerZengqun Zhao, Qingshan LiuACM MM 2021 · 185 citations
- TransFER: Learning Relation-aware Facial Expression Representations with TransformersFanglei Xue, Qiangchang Wang, Guodong GuoICCV 2021 · 276 citations
- Co-Completion for Occluded Facial Expression RecognitionZhen Xing, Weimin Tan, Ruian He, Yangle Lin et al.ACM MM 2022 · 11 citations
- Variance-Aware Bi-Attention Expression Transformer for Open-Set Facial Expression Recognition in the WildJunjie Zhu, Bingjun Luo, Ao Sun, Jinghang Tan et al.ACM MM 2023 · 6 citations
