C2P-CLIP: Injecting Category Common Prompt in CLIP to Enhance Generalization in Deepfake Detection
Chuangchuang Tan, Renshuai Tao, Huan Liu, Guanghua Gu, Baoyuan Wu, Yao Zhao, Yunchao Wei
摘要
This work focuses on AIGC detection to develop universal detectors capable of identifying various types of forgery images. Recent studies have found large pre-trained models, such as CLIP, are effective for generalizable deepfake detection along with linear classifiers. However, two critical issues remain unresolved: 1) understanding why CLIP features are effective on deepfake detection through a linear classifier; and 2) exploring the detection potential of CLIP. In this study, we delve into the underlying mechanisms of CLIP's detection capabilities by decoding its detection features into text and performing word frequency analysis. Our finding indicates that CLIP detects deepfakes by recognizing similar concepts (Fig. 1 a ). Building on this insight, we introduce Category Common Prompt CLIP, called C2P-CLIP, which integrates the category common prompt into the text encoder to inject category-related concepts into the image encoder, thereby enhancing detection performance (Fig. 1 b ). Our method achieves a 12.41% improvement in detection accuracy compared to the original CLIP, without introducing additional parameters during testing. Comprehensive experiments conducted on two widely-used datasets, encompassing 20 generation models, validate the efficacy of the proposed method, demonstrating state-of-the-art performance. The code is available at https://github.com/chuangchuangtan/ C2P-CLIP-DeepfakeDetection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Towards Reliable Identification of Diffusion-based Image ManipulationsAlex Costanzino, Woody Bayliss, Juil Sock, Marc Górriz Blanch 等NeurIPS 2025 · 被引用 4 次
- Attend and Enrich: Enhanced Visual Prompt for Zero-Shot LearningMan Liu, Huihui Bai, Feng Li, Chunjie Zhang 等AAAI 2025 · 被引用 3 次
- Bridging the Gap Between Ideal and Real-World Evaluation: Benchmarking AI-Generated Image Detection in Challenging ScenariosChunxiao Li, Xiaoxiao Wang, Meiling Li, Boming Miao 等ICCV 2025 · 被引用 2 次
- ForgeLens: Data-Efficient Forgery Focus for Generalizable Forgery Image DetectionYingjian Chen, Lei Zhang, Yakun NiuICCV 2025 · 被引用 2 次
- HAMLET-FFD: Hierarchical Adaptive Multi-modal Learning Embeddings Transformation for Face Forgery DetectionJialei Cui, Jianwei Du, Yanzhe Li, Lei Gao 等ACM MM 2025 · 被引用 2 次
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- Rethinking Vision-Language Model in Face Forensics: Multi-Modal Interpretable Forged Face DetectorXiao Guo, Xiufeng Song, Yue Zhang, Xiaohong Liu 等CVPR 2025
- Standing on the Shoulders of Giants: Reprogramming Visual-Language Model for General Deepfake DetectionKaiqing Lin, Yuzhen Lin, Weixiang Li, Taiping Yao 等AAAI 2025 · 被引用 32 次
- A Rich Knowledge Space for Scalable Deepfake DetectionInho Jung, Hyeongjun Choi, Binh Minh Le, Hohyun Na 等ICLR 2026
- Towards More General Video-based Deepfake Detection through Facial Component Guided Adaptation for Foundation ModelYue-Hua Han, Tai-Ming Huang, Kai-Lung Hua, Jun-Cheng ChenCVPR 2025
- PPM-CLIP: Probabilistic Prompt Modeling for Generalizable AI-Generated Image DetectionXinyuan Wang, Yingxin Lai, Zhiming Luo, Zhihui LiuCVPR 2026 · 被引用 1 次
