Forensics Adapter: Adapting CLIP for Generalizable Face Forgery Detection
Xinjie Cui, Yuezun Li, Ao Luo, Jiaran Zhou, Junyu Dong
Abstract
We describe the Forensics Adapter, an adapter network designed to transform CLIP into an effective and generalizable face forgery detector. Although CLIP is highly versatile, adapting it for face forgery detection is nontrivial as forgery-related knowledge is entangled with a wide range of unrelated knowledge. Existing methods treat CLIP merely as a feature extractor, lacking task-specific adaptation, which limits their effectiveness. To address this, we introduce an adapter to learn face forgery traces -the blending boundaries unique to forged faces, guided by task-specific objectives. Then we enhance the CLIP visual tokens with a dedicated interaction strategy that communicates knowledge across CLIP and the adapter. Since the adapter is alongside CLIP, its versatility is highly retained, naturally ensuring strong generalizability in face forgery detection. With only 5.7M trainable parameters, our method achieves a significant performance boost, improving by approximately 7% on average across five standard datasets. We believe the proposed method can serve as a baseline for future CLIP-based face forgery detection methods. The code is available at https://github.com/OUC- VAS/ForensicsAdapter.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- From Specificity to Generality: Revisiting Generalizable Artifacts in Detecting Face DeepfakesLong Ma, Zhiyuan Yan, Jin Xu, Yize Chen et al.NeurIPS 2025 · 24 citations
- WMamba: Wavelet-based Mamba for Face Forgery DetectionSiran Peng, Tianshuo Zhang, Li Gao, Xiangyu Zhu et al.ACM MM 2025 · 17 citations
- NullSwap: Proactive Identity Cloaking Against Deepfake Face SwappingTianyi Wang, Shuaicheng Niu, Harry Cheng, Xiao Zhang et al.ICCV 2025 · 4 citations
- SECOS: Semantic Capture for Rigorous Classification in Open-World Semi-Supervised LearningHezhao Liu, Jiacheng Yang, Junlong Gao, Mengke Li et al.CVPR 2026
- Cross-modal Representation Learning for Diffusion-generated Image DetectionTao Gong, Dayong Wang, Qi Chu, Bin Liu et al.CVPR 2026
Builds on40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- HAMLET-FFD: Hierarchical Adaptive Multi-modal Learning Embeddings Transformation for Face Forgery DetectionJialei Cui, Jianwei Du, Yanzhe Li, Lei Gao et al.ACM MM 2025 · 2 citations
- Towards Generalized Physical Occlusion Detection On DocumentsYiang Zhu, Haoyue Wang, Zhenxing Qian, Sheng Li et al.ACM MM 2025
- FreqBlender: Enhancing DeepFake Detection by Blending Frequency KnowledgeHanzhe Li, Jiaran Zhou, Yuezun Li, Baoyuan Wu et al.NeurIPS 2024 · 96 citations
- ResProto-FD: Visual-Language Residual Prototype Sets for Generalized Face Forgery DetectionJiuyao Jing, Yu Zheng, Chunlei PengAAAI 2026
- Forgery-aware Adaptive Transformer for Generalizable Synthetic Image DetectionHuan Liu, Zichang Tan, Chuangchuang Tan, Yunchao Wei et al.CVPR 2024
