Towards Robust Tampered Text Detection in Document Image: New Dataset and New Solution
Chenfan Qu, Chongyu Liu, Yuliang Liu, Xinhong Chen, Dezhi Peng, Fengjun Guo, Lianwen Jin
Abstract
Recently, tampered text detection in document image has attracted increasingly attention due to its essential role on information security. However, detecting visually consistent tampered text in photographed document images is still a main challenge. In this paper, we propose a novel framework to capture more fine-grained clues in complex scenarios for tampered text detection, termed as Document Tampering Detector (DTD), which consists of a Frequency Perception Head (FPH) to compensate the deficiencies caused by the inconspicuous visual features, and a Multi-view Iterative Decoder (MID) for fully utilizing the information of features in different scales. In addition, we design a new training paradigm, termed as Curriculum Learning for Tampering Detection (CLTD), which can address the confusion during the training procedure and thus to improve the robustness for image compression and the ability to generalize. To further facilitate the tampered text detection in document images, we construct a large-scale document image dataset, termed as DocTamper, which contains 170,000 document images of various types. Experiments demonstrate that our proposed DTD outperforms previous state-of-the-art by 9.2%, 26.3% and 12.3% in terms of F-measure on the DocTamper testing set, and the crossdomain testing sets of DocTamper-FCD and DocTamper-SCD, respectively. Codes and dataset will be available at https://github.com/qcf-568/DocTamper.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- UPOCR: Towards Unified Pixel-Level OCR InterfaceDezhi Peng, Zhenhua Yang, Jiaxin Zhang, Chongyu Liu et al.ICML 2024 · 14 citations
- Omni-IML: Towards Unified Interpretable Image Manipulation LocalizationChenfan Qu, Yiwu Zhong, Fengjun Guo, Lianwen JinICLR 2026 · 5 citations
- ADCD-Net: Robust Document Image Forgery Localization via Adaptive DCT Feature and Hierarchical Content DisentanglementKahim Wong, Jicheng Zhou, Haiwei Wu, Yain-Whar Si et al.ICCV 2025 · 3 citations
- Innovative Image Fraud Detection with Cross-Sample Anomaly Analysis: The Power of LLMsQiwen Wang, Junqi Yang, Zhenghao Lin, Zhenzhe Ying et al.ACL 2025 · 1 citation
- Towards Better Robustness Against Natural Corruptions in Document Tampering LocalizationHuiru Shao, Kaizhu Huang, Wei Wang, Xiaowei Huang et al.AAAI 2025 · 1 citation
Builds on5
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 3,632 citations
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao et al.CVPR 2022 · 2,138 citations
- ObjectFormer for Image Manipulation Detection and LocalizationJunke Wang, Zuxuan Wu, Jingjing Chen, Xintong Han et al.CVPR 2022 · 190 citations
- SwapText: Image Based Texts Transfer in ScenesQiangpeng Yang, Jun Huang, Wei LinCVPR 2020
- STEFANN: Scene Text Editor Using Font Adaptive Neural NetworkPrasun Roy, Saumik Bhattacharya, Subhankar Ghosh, Umapada PalCVPR 2020
Related papers
- Frequency Mining Empowered by Text Aggregation: A New Perspective on Document Image Tampering DetectionZiqi Yi, Guitao Xu, Shihang Wu, Peirong Zhang et al.AAAI 2026
- DITL2: Dual-Stage Invariance Transfer Learning for Generalizable Document Image Tampering LocalizationSongze Li, Yunfei Guo, Shen Chen, Bin Li et al.ACM MM 2025
- From Pixels to Semantics: A Novel MLLM-Driven Approach for Explainable Tampered Text DetectionGuitao Xu, Ziqi Yi, Peirong Zhang, Jiahuan Cao et al.ACM MM 2025 · 2 citations
- D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image DewarpingHeng Li, Xiangping Wu, Qingcai ChenCVPR 2026
- Towards Generalized Physical Occlusion Detection On DocumentsYiang Zhu, Haoyue Wang, Zhenxing Qian, Sheng Li et al.ACM MM 2025
