Suppress and Rebalance: Towards Generalized Multi-Modal Face Anti-Spoofing
Xun Lin, Shuai Wang, Rizhao Cai, Yizhong Liu, Ying Fu, Wenzhong Tang, Zitong Yu, Alex C. Kot
摘要
Face Anti-Spoofing (FAS) is crucial for securing face recognition systems against presentation attacks. With advancements in sensor manufacture and multi-modal learning techniques, many multi-modal FAS approaches have emerged. However, they face challenges in generalizing to unseen attacks and deployment conditions. These challenges arise from (1) modality unreliability, where some modality sensors like depth and infrared undergo significant domain shifts in varying environments, leading to the spread of unreliable information during cross-modal feature fusion, and (2) modality imbalance, where training overly relies on a dominant modality hinders the convergence of others, reducing effectiveness against attack types that are indistinguishable by sorely using the dominant modality. To address modality unreliability, we propose the Uncertainty-Guided Cross-Adapter (U-Adapter) to recognize unreliably detected regions within each modality and suppress the impact of unreliable regions on other modalities. For modality imbalance, we propose a Rebalanced Modality Gradient Modulation (ReGrad) strategy to rebalance the convergence speed of all modalities by adaptively adjusting their gradients. Besides, we provide the first large-scale benchmark for evaluating multi-modal FAS performance under domain generalization scenarios. Extensive experiments demonstrate that our method outperforms state-of-the-art methods. Source codes and protocols are released on https://github.com/OMGGGGG/mmdg .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion RecognitionCam-Van Thi Nguyen, The-Son Le, Anh-Tuan Mai, Duc-Trong LeACM MM 2024 · 被引用 13 次
- Mixture-of-Attack-Experts with Class Regularization for Unified Physical-Digital Face Attack DetectionShunxin Chen, Ajian Liu, Junze Zheng, Jun Wan 等AAAI 2025 · 被引用 11 次
- Multi-View Slot Attention using Paraphrased Texts for Face Anti-SpoofingJeongmin Yu, Susang Kim, Kisu Lee, Taekyoung Kwon 等ICCV 2025 · 被引用 5 次
- TopoTTA: Topology-Enhanced Test-Time Adaptation for Tubular Structure SegmentationJiale Zhou, Wenhan Wang, Shikun Li, Xiaolei Qu 等ICCV 2025 · 被引用 2 次
- DADM: Dual Alignment of Domain and Modality for Face Anti-SpoofingJingyi Yang, Xun Lin, Zitong Yu, Liepiao Zhang 等ICCV 2025 · 被引用 1 次
它引用的顶会 Paper28
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Improving Language Models by Retrieving from Trillions of TokensSebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai 等ICML 2022 · 被引用 1,629 次
- Balanced Multimodal Learning via On-the-fly Gradient ModulationXiaokang Peng, Yake Wei, Andong Deng, Dong Wang 等CVPR 2022 · 被引用 264 次
相关 Paper
- mmFAS: Multimodal Face Anti-Spoofing Using Multi-Level Alignment and Switch-Attention FusionGeng Chen, Wuyuan Xie, Di Lin, Ye Liu 等AAAI 2025 · 被引用 7 次
- Cross Modal Focal Loss for RGBD Face Anti-SpoofingAnjith George, Sébastien MarcelCVPR 2021
- Learning Polysemantic Spoof Trace: A Multi-Modal Disentanglement Network for Face Anti-spoofingKaicheng Li, Hongyu Yang, Binghui Chen, Pengyu Li 等AAAI 2023 · 被引用 4 次
- Adaptive Mixture of Experts Learning for Generalizable Face Anti-SpoofingQianyu Zhou, Ke-Yue Zhang, Taiping Yao, Ran Yi 等ACM MM 2022 · 被引用 65 次
- Towards Unsupervised Domain Generalization for Face Anti-SpoofingYuchen Liu, Yabo Chen, Mengran Gou, Chun-Ting Huang 等ICCV 2023 · 被引用 41 次
