ADCD-Net: Robust Document Image Forgery Localization via Adaptive DCT Feature and Hierarchical Content Disentanglement
Kahim Wong, Jicheng Zhou, Haiwei Wu, Yain-Whar Si, Jiantao Zhou
Abstract
The advancement of image editing tools has enabled malicious manipulation of sensitive document images, underscoring the need for robust document image forgery detection.Though forgery detectors for natural images have been extensively studied, they struggle with document images, as the tampered regions can be seamlessly blended into the uniform document background (BG) and structured text. On the other hand, existing document-specific methods lack sufficient robustness against various degradations, which limits their practical deployment. This paper presents ADCD-Net, a robust document forgery localization model that adaptively leverages the RGB/DCT forensic traces and integrates key characteristics of document images. Specifically, to address the DCT traces' sensitivity to block misalignment, we adaptively modulate the DCT feature contribution based on a predicted alignment score, resulting in much improved resilience to various distortions, including resizing and cropping. Also, a hierarchical content disentanglement approach is proposed to boost the localization performance via mitigating the text-BG disparities. Furthermore, noticing the predominantly pristine nature of BG regions, we construct a pristine prototype capturing traces of untampered regions, and eventually enhance both the localization accuracy and robustness. Our proposed ADCD-Net demonstrates superior forgery localization performance, consistently outperforming state-of-the-art methods by 20.79% averaged over 5 types of distortions. The code is available at https://github.com/KAHIMWONG/ACDC-Net.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b3209f26-1d70-4b51-8e12-a9daedc96a9aCited by top-tier papers4
- TextShield-R1: Reinforced Reasoning for Tampered Text DetectionChenfan Qu, Yiwu Zhong, Jian Liu, Xuekang Zhu et al.AAAI 2026 · 4 citations
- Editprint: General Digital Image Forensics via Editing Fingerprint with Self-Augmentation TrainingHaiwei Wu, Kemou Li, Yuanman Li, Jiantao ZhouCVPR 2026
- Detect Any AI-Counterfeited Text ImageChenfan Qu, Yiwu Zhong, Xuekang Zhu, Junchi Li et al.CVPR 2026
- Zero-shot Detection of AI-Generated Image via RAW-RGB AlignmentHaiwei Wu, Fengpeng Li, Zhilin Tu, Yuanman Li et al.CVPR 2026
Builds on14
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- UCF: Uncovering Common Features for Generalizable Deepfake DetectionZhiyuan Yan, Yong Zhang, Yanbo Fan, Baoyuan WuICCV 2023 · 264 citations
Related papers
- Towards Generalized Physical Occlusion Detection On DocumentsYiang Zhu, Haoyue Wang, Zhenxing Qian, Sheng Li et al.ACM MM 2025
- Frequency Mining Empowered by Text Aggregation: A New Perspective on Document Image Tampering DetectionZiqi Yi, Guitao Xu, Shihang Wu, Peirong Zhang et al.AAAI 2026
- Learning Discriminative Noise Guidance for Image Forgery Detection and LocalizationJiaying Zhu, Dong Li, Xueyang Fu, Gang Yang et al.AAAI 2024 · 27 citations
- Self-supervised Domain Adaptation for Forgery Localization of JPEG Compressed ImagesYuan Rao, Jiangqun NiICCV 2021 · 34 citations
- Forensics Adapter: Adapting CLIP for Generalizable Face Forgery DetectionXinjie Cui, Yuezun Li, Ao Luo, Jiaran Zhou et al.CVPR 2025
