Detecting Synthetic Image by Cross-Modal Commonality Interaction
Kai Li, Wenqi Ren, Wei Wang, Linchao Zhang, Xiaochun Cao
摘要
Existing synthetic image detection approaches can be categorized into three paradigms: spatial, frequency, and fingerprint-based methods. Our analysis reveals a fundamental commonality across these paradigms: a significant reliance on high-frequency image components. This observation highlights the discriminative power of high-frequency information for this task and provides a strong rationale for learning generalized artifact representations based on multi-modal fusion strategies. Building on this insight, we introduce a multi-modal high-frequency interactive detection framework for general synthetic image detection. This framework explicitly integrates high-frequency information from both the spatial and frequency domains. Specifically, its spatial processing branch incorporates a novel high-frequency self-enhancement module to bolster local high-frequency representations. Concurrently, the frequency processing branch utilizes a multi-scale frequency information enhancement module to capture diverse contextual cues. At the feature fusion stage, we propose a pooling-guided cross-modal high-frequency interaction module, which dynamically weights cross-modal information to further reinforce salient high-frequency representations. Extensive experiments on public datasets demonstrate that our proposed framework achieves state-of-the-art performance in real-world detection scenarios.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Generalizing Face Forgery Detection With High-Frequency FeaturesYuchen Luo, Yong Zhang, Junchi Yan, Wei LiuCVPR 2021
- Frequency-Aware Deepfake Detection: Improving Generalizability through Frequency Space Domain LearningChuangchuang Tan, Yao Zhao, Shikui Wei, Guanghua Gu 等AAAI 2024 · 被引用 232 次
- SynerDetect: Hierarchical Synergistic Learning for Generalizable AI-Generated Image DetectionShuaibo Li, Yijun Yang, Zhaohu Xing, Hongqiu Wang 等AAAI 2026
- Knowledge-Enhanced Multimodal Fake News Detection: Semantic Visual and Priority FusionQin Zhang, Jiaying Liu, Qian Tao, Zhiwei Guo 等WWW 2026
- Probing Synergistic High-Order Interaction in Infrared and Visible Image FusionNaishan Zheng, Man Zhou, Jie Huang, Junming Hou 等CVPR 2024 · 被引用 43 次
