Towards Generalizable Detector for Generated Image
Qianshu Cai, Chao Wu, Yonggang Zhang, Jun Yu, Xinmei Tian
Abstract
The effective detection of generated images is crucial to mitigate potential risks associated with their misuse. Despite significant progress, a fundamental challenge remains: ensuring the generalizability of detectors. To address this, we propose a novel perspective on understanding and improving generated image detection, inspired by the human cognitive process: Humans identify an image as unnatural based on specific patterns because these patterns lie outside the space spanned by those of natural images. This is intrinsically related to out-of-distribution (OOD) detection, which identifies samples whose semantic patterns (i.e., labels) lie outside the semantic pattern space of in-distribution (ID) samples. By treating patterns of generated images as OOD samples, we demonstrate that models trained merely over natural images bring guaranteed generalization ability under mild assumptions. This transforms the generalization challenge of generated image detection into the problem of fitting natural image patterns. Based on this insight, we propose a generalizable detection method through the lens of ID energy. Theoretical results capture the generalization risk of the proposed method. Experimental results across multiple benchmarks demonstrate the effectiveness of our approach. Code is available at https://github.com/dav-joy-thon/DEnD-Detection .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fe50aaeb-d45f-447d-bb5f-ea12561bf6beCited by top-tier papers2
- GenShield: Unified Detection and Artifact Correction for AI-Generated ImagesZhipei Xu, Xuanyu Zhang, Youmin Xu, Qing Huang et al.ICML 2026 · 1 citation
- Detect Any AI-Counterfeited Text ImageChenfan Qu, Yiwu Zhong, Xuekang Zhu, Junchi Li et al.CVPR 2026
Builds on36
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Beyond Semantic Features: Pixel-level Mapping for Generalized AI-Generated Image DetectionChenming Zhou, Jiaan Wang, Yu Li, Lei Li et al.AAAI 2026 · 1 citation
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 2,213 citations
- DiffGuard: Semantic Mismatch-Guided Out-of-Distribution Detection using Pre-trained Diffusion ModelsRuiyuan Gao, Chenchen Zhao, Lanqing Hong, Qiang XuICCV 2023 · 29 citations
- Learning by Erasing: Conditional Entropy Based Transferable Out-of-Distribution DetectionMeng Xing, Zhiyong Feng, Yong Su, Changjae OhAAAI 2024 · 7 citations
- Diffusion-based Semantic Outlier Generation via Nuisance Awareness for Out-of-Distribution DetectionSuhee Yoon, Sanghyu Yoon, Ye Seul Sim, Sungik Choi et al.AAAI 2025 · 3 citations
