PiD: Generalized AI-Generated Images Detection with Pixelwise Decomposition Residuals
Xinghe Fu, Zhiyuan Yan, Zheng Yang, Taiping Yao, Yandan Zhao, Shouhong Ding, Xi Li
摘要
Fake images, created by recently advanced generative models, have become increasingly indistinguishable from real ones, making their detection crucial, urgent, and challenging. This paper introduces PiD (Pixelwise Decomposition Residuals), a novel detection method that focuses on residual signals within images. Generative models are designed to optimize high-level semantic content (principal components), often overlooking lowlevel signals (residual components). PiD leverages this observation by disentangling residual components from images, encouraging the model to uncover more underlying and general forgery clues independent of semantic content. Compared to prior approaches that rely on reconstruction techniques or high-frequency information, PiD is computationally efficient and does not rely on any generative models for reconstruction. Specifically, PiD operates at the pixel level, mapping the pixel vector to another color space (e.g., YUV) and then quantizing the vector. The pixel vector is mapped back to the RGB space and the quantization loss is taken as the residual for AIGC detection. Our experiment results are striking and highly surprising: PiD achieves 98% accuracy on the widely used GenImage benchmark, highlighting the effectiveness and generalization performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- FiSeR: Fine-Grained Source Representations for Cross-Domain AI Image DetectionShan Zhang, Yongxin He, Mingming Zhang, Huiwen Tian 等ICML 2026
- Dissect and Prune: Enhancing Robustness in AI-Generated Image DetectionDahye Kim, Jaehyun Choi, Hyun Seok Seong, Seongho Kim 等ICML 2026
- Your AI-Generated Image Detector Can Secretly Achieve SOTA Accuracy, If CalibratedMuli Yang, Gabriel James Goenawan, Henan Wang, Huaiyuan Qin 等AAAI 2026
它引用的顶会 Paper30
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- STD-FD: Spatio-Temporal Distribution Fitting Deviation for AIGC Forgery IdentificationHengrui Lou, Zunlei Feng, Jinsong Geng, Erteng Liu 等ICML 2025
- Spatial-Temporal Forgery Trace Based Forgery Image IdentificationYilin Wang, Zunlei Feng, Jiachi Wang, Hengrui Lou 等ICCV 2025 · 被引用 1 次
- Attribution as Retrieval: Model-Agnostic AI-Generated Image AttributionHongsong Wang, Renxi Cheng, Chaolei Han, Jie GuiCVPR 2026 · 被引用 4 次
- SemGIR: Semantic-Guided Image Regeneration Based Method for AI-generated Image Detection and AttributionXiao Yu, Kejiang Chen, Kai Zeng, Han Fang 等ACM MM 2024 · 被引用 8 次
- FIND: A Simple Yet Effective Baseline for Diffusion-Generated Image DetectionJie Li, Yingying Feng, Chi Xie, Jie Hu 等AAAI 2026
