Decision-Driven Orthogonal Learning with Complementary Feature Mining for Robust Synthetic Image Detection
Kai Li, Wei Wang, Linchao Zhang, Siying Zhu, Wenqi Ren
Abstract
The widespread and inconsistent compression applied by Online Social Networks severely degrades the performance of synthetic image detectors. We attribute this degradation to two main issues: 1) the model confuses forgery artifacts with compression artifacts, and 2) compression erodes crucial discriminative high-frequency details. Existing methods suppress compression features during training but overlook the overlap between compression features and forgery-related features, leading to the unintended removal of forgery traces. To address artifact confusion, we introduce a Decision-Driven Orthogonal Constraint, which defines a classification decision axis pointing from the real class centroid to the forged class centroid. This constraint enforces compression artifacts to be orthogonal to the decision axis, mitigating their interference with forgery detection without entirely removing them, thus preventing the suppression of forgery-related features. To mitigate the erosion of high-frequency details, we propose to mine complementary forgery cues from both low-frequency information and compressed high-frequency components. A bidirectional update strategy and an adaptive global-local modulator are proposed to facilitate the utilization of forgery cues. Extensive experiments demonstrate that our method achieves state-of-the-art generalization performance in challenging open-world detection scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8fb4bed9-8d99-4c74-9144-d6e29b360789Builds on22
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- Leveraging Frequency Analysis for Deep Fake Image RecognitionJoel Frank, Thorsten Eisenhofer, Lea Schönherr, Asja Fischer et al.ICML 2020 · 848 citations
- How Do Vision Transformers Work?Namuk Park, Songkuk KimICLR 2022 · 653 citations
Related papers
- Metric Learning for Anti-Compression Facial Forgery DetectionShenhao Cao, Qin Zou, Xiuqing Mao, Dengpan Ye et al.ACM MM 2021 · 23 citations
- Towards Open-world Generalized Deepfake Detection: General Feature Extraction via Unsupervised Domain AdaptationMidou Guo, Qilin Yin, Wei Lu, Xiangyang LuoACM MM 2025 · 2 citations
- ODDN: Addressing Unpaired Data Challenges in Open-World Deepfake Detection on Online Social NetworksRenshuai Tao, Manyi Le, Chuangchuang Tan, Huan Liu et al.AAAI 2025 · 7 citations
- End-to-End Reconstruction-Classification Learning for Face Forgery DetectionJunyi Cao, Chao Ma, Taiping Yao, Shen Chen et al.CVPR 2022 · 327 citations
- JPEG Compression-aware Image Forgery LocalizationMenglu Wang, Xueyang Fu, Jiawei Liu, Zheng-Jun ZhaACM MM 2022 · 22 citations
