Amplifying Discrepancies: Exploiting Macro and Micro Inconsistencies for Image Manipulation Localization
Shenghao Chen, Yibo Zhao, Tianyi Wang, Chunjie Ma, Weili Guan, Ming Li, Zan Gao
Abstract
The rapid development of image manipulation technologies poses significant challenges to multimedia forensics, especially in accurate localization of manipulated regions. Existing methods often fail to fully explore the intrinsic discrepancies between manipulated and authentic regions, resulting in sub-optimal performance. To address this limitation, we propose the Focus Region Discrepancy Network (FRD-Net), a novel and efficient framework that significantly enhances manipulation localization by amplifying discrepancies at both macro- and micro-levels. Specifically, our proposed Iterative Clustering Module (ICM) groups features into two discriminative clusters and refines representations via backward propagation from cluster centers, improving the distinction between tampered and authentic regions at the macro level. Thereafter, our Differential Progressive Module (DPM) is constructed to capture fine-grained structural inconsistencies within local neighborhoods and integrate them into a Central Difference Convolution, increasing sensitivity to subtle manipulation details at the micro level. Finally, these complementary modules are seamlessly integrated into a compact architecture that achieves a favorable balance between accuracy and efficiency. Extensive experiments on multiple benchmarks demonstrate that FRD-Net consistently surpasses state-of-the-art methods in terms of manipulation localization performance while maintaining a lower computational cost.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on15
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam et al.ICML 2022 · 4,691 citations
- Anomaly Transformer: Time Series Anomaly Detection with Association DiscrepancyJiehui Xu, Haixu Wu, Jianmin Wang, Mingsheng LongICLR 2022 · 960 citations
- Pixel Difference Networks for Efficient Edge DetectionZhuo Su, Wenzhe Liu, Zitong Yu, Dewen Hu et al.ICCV 2021 · 488 citations
Related papers
- DiffForensics: Leveraging Diffusion Prior to Image Forgery Detection and LocalizationZeqin Yu, Jiangqun Ni, Yuzhen Lin, Haoyi Deng et al.CVPR 2024 · 25 citations
- CatmullRom Splines-Based Regression for Image Forgery LocalizationLi Zhang, Mingliang Xu, Dong Li, Jianming Du et al.AAAI 2024 · 9 citations
- M²RL-Net: Multi-View and Multi-Level Relation Learning Network for Weakly-Supervised Image Forgery DetectionJiafeng Li, Ying Wen, Lianghua HeAAAI 2025 · 2 citations
- A New Benchmark and Model for Challenging Image Manipulation DetectionZhenfei Zhang, Mingyang Li, Ming-Ching ChangAAAI 2024 · 18 citations
- Collaborative Transformers with Multi-Level Forensic Attention for Image Manipulation LocalizationJiwei Zhang, Wenbo Feng, Siwei Wang, Feifei Kou et al.AAAI 2026
