Detecting Compressed AI-Generated Images via Phase Spectrum Robustness
Kai Li, Wenqi Ren, Wei Wang, Xiaochun Cao
摘要
This paper aims to present a robust AI-generated image detection framework designed to address performance degradation caused by image compression in online social networks. The key challenges are twofold: 1) compression destroys fragile artifacts that are crucial to existing methods, and 2) it introduces new compression artifacts that interfere with detection. Existing methods typically enhance the compression robustness by collecting original-compression pairs and compression labels. However, the collection and annotation process is highly resource-intensive. To address these issues, we propose a Compression-Robust Phase-Harmonized Transformer, motivated by the observation that phase spectrum remains stable under compression. The framework consists of a phase-harmonized cross-modal interaction module that leverages phase spectrum information for feature fusion, enhancing compression robustness, and a multi-domain modulation adapter that further refines fused features while enabling parameter-efficient finetuning. In particular, the framework operates without requiring compression-original data pairs and compression labels. When limited compression labels are available, we introduce a difficulty-aware consistency loss to maximize their utility by prioritizing hard compressed samples during training, further boosting robustness. Extensive experiments demonstrate that our method significantly outperforms state-of-the-art approaches, exhibiting superior robustness against image compression.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper25
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess 等ICCV 2019 · 被引用 2,966 次
相关 Paper
- Metric Learning for Anti-Compression Facial Forgery DetectionShenhao Cao, Qin Zou, Xiuqing Mao, Dengpan Ye 等ACM MM 2021 · 被引用 23 次
- Decision-Driven Orthogonal Learning with Complementary Feature Mining for Robust Synthetic Image DetectionKai Li, Wei Wang, Linchao Zhang, Siying Zhu 等AAAI 2026
- Dissect and Prune: Enhancing Robustness in AI-Generated Image DetectionDahye Kim, Jaehyun Choi, Hyun Seok Seong, Seongho Kim 等ICML 2026
- ODDN: Addressing Unpaired Data Challenges in Open-World Deepfake Detection on Online Social NetworksRenshuai Tao, Manyi Le, Chuangchuang Tan, Huan Liu 等AAAI 2025 · 被引用 7 次
- Beyond Pixels: Mining Compressed Domain Artifacts for Efficient AI-Generated Video DetectionAnran Zhu, Zhengli Shi, Chende Zheng, Chenhao Lin 等ICML 2026
