Suppress and Rebalance: Towards Generalized Multi-Modal Face Anti-Spoofing
Xun Lin, Shuai Wang, Rizhao Cai, Yizhong Liu, Ying Fu, Wenzhong Tang, Zitong Yu, Alex C. Kot
Abstract
Face Anti-Spoofing (FAS) is crucial for securing face recognition systems against presentation attacks. With advancements in sensor manufacture and multi-modal learning techniques, many multi-modal FAS approaches have emerged. However, they face challenges in generalizing to unseen attacks and deployment conditions. These challenges arise from (1) modality unreliability, where some modality sensors like depth and infrared undergo significant domain shifts in varying environments, leading to the spread of unreliable information during cross-modal feature fusion, and (2) modality imbalance, where training overly relies on a dominant modality hinders the convergence of others, reducing effectiveness against attack types that are indistinguishable by sorely using the dominant modality. To address modality unreliability, we propose the Uncertainty-Guided Cross-Adapter (U-Adapter) to recognize unreliably detected regions within each modality and suppress the impact of unreliable regions on other modalities. For modality imbalance, we propose a Rebalanced Modality Gradient Modulation (ReGrad) strategy to rebalance the convergence speed of all modalities by adaptively adjusting their gradients. Besides, we provide the first large-scale benchmark for evaluating multi-modal FAS performance under domain generalization scenarios. Extensive experiments demonstrate that our method outperforms state-of-the-art methods. Source codes and protocols are released on https://github.com/OMGGGGG/mmdg .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 18a1a4c1-c60e-40b1-878b-330574e8f649Cited by top-tier papers10
- Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion RecognitionCam-Van Thi Nguyen, The-Son Le, Anh-Tuan Mai, Duc-Trong LeACM MM 2024 · 13 citations
- Mixture-of-Attack-Experts with Class Regularization for Unified Physical-Digital Face Attack DetectionShunxin Chen, Ajian Liu, Junze Zheng, Jun Wan et al.AAAI 2025 · 11 citations
- Multi-View Slot Attention using Paraphrased Texts for Face Anti-SpoofingJeongmin Yu, Susang Kim, Kisu Lee, Taekyoung Kwon et al.ICCV 2025 · 5 citations
- TopoTTA: Topology-Enhanced Test-Time Adaptation for Tubular Structure SegmentationJiale Zhou, Wenhan Wang, Shikun Li, Xiaolei Qu et al.ICCV 2025 · 2 citations
- DADM: Dual Alignment of Domain and Modality for Face Anti-SpoofingJingyi Yang, Xun Lin, Zitong Yu, Liepiao Zhang et al.ICCV 2025 · 1 citation
Builds on28
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- Improving Language Models by Retrieving from Trillions of TokensSebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai et al.ICML 2022 · 1,629 citations
- Balanced Multimodal Learning via On-the-fly Gradient ModulationXiaokang Peng, Yake Wei, Andong Deng, Dong Wang et al.CVPR 2022 · 264 citations
Related papers
- mmFAS: Multimodal Face Anti-Spoofing Using Multi-Level Alignment and Switch-Attention FusionGeng Chen, Wuyuan Xie, Di Lin, Ye Liu et al.AAAI 2025 · 7 citations
- Cross Modal Focal Loss for RGBD Face Anti-SpoofingAnjith George, Sébastien MarcelCVPR 2021
- Learning Polysemantic Spoof Trace: A Multi-Modal Disentanglement Network for Face Anti-spoofingKaicheng Li, Hongyu Yang, Binghui Chen, Pengyu Li et al.AAAI 2023 · 4 citations
- Adaptive Mixture of Experts Learning for Generalizable Face Anti-SpoofingQianyu Zhou, Ke-Yue Zhang, Taiping Yao, Ran Yi et al.ACM MM 2022 · 65 citations
- Towards Unsupervised Domain Generalization for Face Anti-SpoofingYuchen Liu, Yabo Chen, Mengran Gou, Chun-Ting Huang et al.ICCV 2023 · 41 citations
