Dual Data Alignment Makes AI-Generated Image Detector Easier Generalizable
Ruoxin Chen, Junwei Xi, Zhiyuan Yan, Ke-Yue Zhang, Shuang Wu, Jingyi Xie, Xu Chen, Lei Xu, Isabel Guan, Taiping Yao, Shouhong Ding
摘要
Existing detectors are often trained on biased datasets, leading to the possibility of overfitting on non-causal image attributes that are spuriously correlated with real/synthetic labels. While these biased features enhance performance on the training data, they result in substantial performance degradation when applied to unbiased datasets. One common solution is to perform dataset alignment through generative reconstruction, matching the semantic content between real and synthetic images. However, we revisit this approach and show that pixel-level alignment alone is insufficient. The reconstructed images still suffer from frequency-level misalignment, which can perpetuate spurious correlations. To illustrate, we observe that reconstruction models tend to restore the high-frequency details lost in real images (possibly due to JPEG compression), inadvertently creating a frequency-level misalignment, where synthetic images appear to have richer high-frequency content than real ones. This misalignment leads to models associating high-frequency features with synthetic labels, further reinforcing biased cues. To resolve this, we propose Dual Data Alignment (DDA), which aligns both the pixel and frequency domains. Moreover, we introduce two new test sets: DDA-COCO, containing DDA-aligned synthetic images for testing detector performance on the most aligned dataset, and EvalGEN, featuring the latest generative models for assessing detectors under new generative architectures such as visual auto-regressive generators. Finally, our extensive evaluations demonstrate that a detector trained exclusively on DDA-aligned MSCOCO could improve across 8 diverse benchmarks by a non-trivial margin, showing a +7.2% on in-the-wild benchmarks, highlighting the improved generalizability of unbiased detectors. Our code is available at: https://github.com/roy-ch/Dual-Data-Alignment.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Spot the Fake: Large Multimodal Model-Based Synthetic Image Detection with Artifact ExplanationSiwei Wen, Junyan Ye, Peilin Feng, Hengrui Kang 等NeurIPS 2025 · 被引用 82 次
- X2-DFD: A framework for explainable and extendable Deepfake DetectionYize Chen, Zhiyuan Yan, Guangliang Cheng, Kangran Zhao 等NeurIPS 2025 · 被引用 43 次
- Breaking Latent Prior Bias in Detectors for Generalizable AIGC Image DetectionYue Zhou, Xinan He, Kaiqing Lin, Bing Fan 等NeurIPS 2025 · 被引用 29 次
- All Patches Matter, More Patches Better: Enhance AI-Generated Image Detection via Panoptic Patch LearningZheng Yang, Ruoxin Chen, Zhiyuan Yan, Ke-Yue Zhang 等ICLR 2026 · 被引用 28 次
- VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement LearningHao Tan, jun lan, Senyuan Shi, Zichang Tan 等ICML 2026 · 被引用 12 次
它引用的顶会 Paper30
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- ControlVideo: Training-free Controllable Text-to-video GenerationYabo Zhang, Yuxiang Wei, Dongsheng Jiang, Xiaopeng Zhang 等ICLR 2024 · 被引用 359 次
- Frequency-Aware Deepfake Detection: Improving Generalizability through Frequency Space Domain LearningChuangchuang Tan, Yao Zhao, Shikui Wei, Guanghua Gu 等AAAI 2024 · 被引用 232 次
- Rethinking the Up-Sampling Operations in CNN-Based Generative Network for Generalizable Deepfake DetectionChuangchuang Tan, Huan Liu, Yao Zhao, Shikui Wei 等CVPR 2024 · 被引用 126 次
相关 Paper
- SONAR: Spectral‑Contrastive Audio Residuals for Generalizable Deepfake DetectionIdo Nitzan Hidekel, Gal Lifshitz, Khen Cohen, Dan RavivICML 2026 · 被引用 1 次
- Aggregating Diverse Cue Experts for AI-Generated Image DetectionLei Tan, Shuwei Li, Mohan Kankanhalli, Robby T. TanAAAI 2026
- Boosting Domain Generalized and Adaptive Detection with Diffusion Models: Fitness, Generalization, and TransferabilityBoyong He, Yuxiang Ji, Zhuoyue Tan, Liaoni WuICCV 2025 · 被引用 3 次
- IGG: Improved Graph Generation for Domain Adaptive Object DetectionPengteng Li, Ying He, F. Richard Yu, Pinhao Song 等ACM MM 2023 · 被引用 10 次
- UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image DetectionYanran Zhang, Wenzhao Zheng, Yifei Li, Bingyao Yu 等CVPR 2026 · 被引用 3 次
