Towards Modern Image Manipulation Localization: A Large-Scale Dataset and Novel Methods
Chenfan Qu, Yiwu Zhong, Chongyu Liu, Guitao Xu, Dezhi Peng, Fengjun Guo, Lianwen Jin
摘要
In recent years, image manipulation localization has attracted increasing attention due to its pivotal role in guaranteeing social media security. However, how to accurately identify the forged regions remains an open challenge. One of the main bottlenecks lies in the severe scarcity of high-quality data, due to its costly creation process. To address this limitation, we propose a novel paradigm, termed as CAAA, to automatically and precisely annotate the numerous manually forged images from the web at the pixel level. We further propose a novel metric QES to facilitate the automatic filtering of unreliable annotations. With CAAA and QES, we construct a large-scale, diverse, and high-quality dataset comprising 123,150 manually forged images with mask annotations. Besides, we develop a new model APSC-Net for accurate image manipulation localization. According to extensive experiments, our dataset significantly improves the performance of various models on the widely-used benchmarks and such improvements are attributed to our proposed effective methods. The dataset and code are publicly available at https://github.com/qcf-568/MIML.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- ForgerySleuth: Empowering Multimodal Large Language Models for Image Manipulation DetectionZhihao Sun, Haoran Jiang, Haoran Chen, Yixin Cao 等NeurIPS 2025 · 被引用 16 次
- Omni-IML: Towards Unified Interpretable Image Manipulation LocalizationChenfan Qu, Yiwu Zhong, Fengjun Guo, Lianwen JinICLR 2026 · 被引用 5 次
- Towards Reliable Identification of Diffusion-based Image ManipulationsAlex Costanzino, Woody Bayliss, Juil Sock, Marc Górriz Blanch 等NeurIPS 2025 · 被引用 4 次
- TextShield-R1: Reinforced Reasoning for Tampered Text DetectionChenfan Qu, Yiwu Zhong, Jian Liu, Xuekang Zhu 等AAAI 2026 · 被引用 4 次
- ADCD-Net: Robust Document Image Forgery Localization via Adaptive DCT Feature and Hierarchical Content DisentanglementKahim Wong, Jicheng Zhou, Haiwei Wu, Yain-Whar Si 等ICCV 2025 · 被引用 3 次
它引用的顶会 Paper11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- ObjectFormer for Image Manipulation Detection and LocalizationJunke Wang, Zuxuan Wu, Jingjing Chen, Xintong Han 等CVPR 2022 · 被引用 190 次
- Robust Image Forgery Detection over Online Social Network Shared ImagesHaiwei Wu, Jiantao Zhou, Jinyu Tian, Jun LiuCVPR 2022 · 被引用 78 次
- Pre-training-free Image Manipulation Localization through Non-Mutually Exclusive Contrastive LearningJizhe Zhou, Xiaochen Ma, Xia Du, Ahmed Y. Al Hammadi 等ICCV 2023 · 被引用 51 次
相关 Paper
- Face Forensics in the WildTianfei Zhou, Wenguan Wang, Zhiyuan Liang, Jianbing ShenCVPR 2021
- SIDA: Social Media Image Deepfake Detection, Localization and Explanation with Large Multimodal ModelZhenglin Huang, Jinwei Hu, Xiangtai Li, Yiwei He 等CVPR 2025
- Generic Image Manipulation Localization through the Lens of Multi-scale Spatial InconsistenceZan Gao, Shenghao Chen, Yangyang Guo, Weili Guan 等ACM MM 2022 · 被引用 14 次
- On the Detection of Digital Face ManipulationHao Dang, Feng Liu, Joel Stehouwer, Xiaoming Liu 等CVPR 2020
- AV-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake DatasetZhixi Cai, Shreya Ghosh, Aman Pankaj Adatia, Munawar Hayat 等ACM MM 2024 · 被引用 51 次
