Cross Modal Focal Loss for RGBD Face Anti-Spoofing
Anjith George, Sébastien Marcel
Abstract
Automatic methods for detecting presentation attacks are essential to ensure the reliable use of facial recognition technology. Most of the methods available in the literature for presentation attack detection (PAD) fails in generalizing to unseen attacks. In recent years, multi-channel methods have been proposed to improve the robustness of PAD systems. Often, only a limited amount of data is available for additional channels, which limits the effectiveness of these methods. In this work, we present a new framework for PAD that uses RGB and depth channels together with a novel loss function. The new architecture uses complementary information from the two modalities while reducing the impact of overfitting. Essentially, a cross-modal focal loss function is proposed to modulate the loss contribution of each channel as a function of the confidence of individual channels. Extensive evaluations in two publicly available datasets demonstrate the effectiveness of the proposed approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ccfc4e9f-c463-445f-8eaf-49f688f6f764Cited by top-tier papers16
- DFormer: Rethinking RGBD Representation Learning for Semantic SegmentationBowen Yin, Xuying Zhang, Zhong-Yu Li, Li Liu et al.ICLR 2024 · 110 citations
- Interact, Embed, and EnlargE: Boosting Modality-Specific Representations for Multi-Modal Person Re-identificationZi Wang, Chenglong Li, Aihua Zheng, Ran He et al.AAAI 2022 · 61 citations
- CFPL-FAS: Class Free Prompt Learning for Generalizable Face Anti-SpoofingAjian Liu, Shuai Xue, Jianwen Gan, Jun Wan et al.CVPR 2024 · 59 citations
- FM-CLIP: Flexible Modal CLIP for Face Anti-SpoofingAjian Liu, Hui Ma, Junze Zheng, Haocheng Yuan et al.ACM MM 2024 · 34 citations
- Causal Inference over Visual-Semantic-Aligned Graph for Image ClassificationLei Meng, Xiangxian Li, Xiaoshuo Yan, Haokai Ma et al.AAAI 2025 · 11 citations
Related papers
- Suppress and Rebalance: Towards Generalized Multi-Modal Face Anti-SpoofingXun Lin, Shuai Wang, Rizhao Cai, Yizhong Liu et al.CVPR 2024
- mmFAS: Multimodal Face Anti-Spoofing Using Multi-Level Alignment and Switch-Attention FusionGeng Chen, Wuyuan Xie, Di Lin, Ye Liu et al.AAAI 2025 · 7 citations
- Deep Spatial Gradient and Temporal Depth Learning for Face Anti-SpoofingZezheng Wang, Zitong Yu, Chenxu Zhao, Xiangyu Zhu et al.CVPR 2020
- Cross-Domain Face Presentation Attack Detection via Multi-Domain Disentangled Representation LearningGuoqing Wang, Hu Han, Shiguang Shan, Xilin ChenCVPR 2020
- Learning Polysemantic Spoof Trace: A Multi-Modal Disentanglement Network for Face Anti-spoofingKaicheng Li, Hongyu Yang, Binghui Chen, Pengyu Li et al.AAAI 2023 · 4 citations
