Towards Robust Tampered Text Detection in Document Image: New Dataset and New Solution
Chenfan Qu, Chongyu Liu, Yuliang Liu, Xinhong Chen, Dezhi Peng, Fengjun Guo, Lianwen Jin
摘要
Recently, tampered text detection in document image has attracted increasingly attention due to its essential role on information security. However, detecting visually consistent tampered text in photographed document images is still a main challenge. In this paper, we propose a novel framework to capture more fine-grained clues in complex scenarios for tampered text detection, termed as Document Tampering Detector (DTD), which consists of a Frequency Perception Head (FPH) to compensate the deficiencies caused by the inconspicuous visual features, and a Multi-view Iterative Decoder (MID) for fully utilizing the information of features in different scales. In addition, we design a new training paradigm, termed as Curriculum Learning for Tampering Detection (CLTD), which can address the confusion during the training procedure and thus to improve the robustness for image compression and the ability to generalize. To further facilitate the tampered text detection in document images, we construct a large-scale document image dataset, termed as DocTamper, which contains 170,000 document images of various types. Experiments demonstrate that our proposed DTD outperforms previous state-of-the-art by 9.2%, 26.3% and 12.3% in terms of F-measure on the DocTamper testing set, and the crossdomain testing sets of DocTamper-FCD and DocTamper-SCD, respectively. Codes and dataset will be available at https://github.com/qcf-568/DocTamper.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- UPOCR: Towards Unified Pixel-Level OCR InterfaceDezhi Peng, Zhenhua Yang, Jiaxin Zhang, Chongyu Liu 等ICML 2024 · 被引用 14 次
- Omni-IML: Towards Unified Interpretable Image Manipulation LocalizationChenfan Qu, Yiwu Zhong, Fengjun Guo, Lianwen JinICLR 2026 · 被引用 5 次
- ADCD-Net: Robust Document Image Forgery Localization via Adaptive DCT Feature and Hierarchical Content DisentanglementKahim Wong, Jicheng Zhou, Haiwei Wu, Yain-Whar Si 等ICCV 2025 · 被引用 3 次
- Innovative Image Fraud Detection with Cross-Sample Anomaly Analysis: The Power of LLMsQiwen Wang, Junqi Yang, Zhenghao Lin, Zhenzhe Ying 等ACL 2025 · 被引用 1 次
- Towards Better Robustness Against Natural Corruptions in Document Tampering LocalizationHuiru Shao, Kaizhu Huang, Wei Wang, Xiaowei Huang 等AAAI 2025 · 被引用 1 次
它引用的顶会 Paper5
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao 等CVPR 2022 · 被引用 2,138 次
- ObjectFormer for Image Manipulation Detection and LocalizationJunke Wang, Zuxuan Wu, Jingjing Chen, Xintong Han 等CVPR 2022 · 被引用 190 次
- SwapText: Image Based Texts Transfer in ScenesQiangpeng Yang, Jun Huang, Wei LinCVPR 2020
- STEFANN: Scene Text Editor Using Font Adaptive Neural NetworkPrasun Roy, Saumik Bhattacharya, Subhankar Ghosh, Umapada PalCVPR 2020
相关 Paper
- Frequency Mining Empowered by Text Aggregation: A New Perspective on Document Image Tampering DetectionZiqi Yi, Guitao Xu, Shihang Wu, Peirong Zhang 等AAAI 2026
- DITL2: Dual-Stage Invariance Transfer Learning for Generalizable Document Image Tampering LocalizationSongze Li, Yunfei Guo, Shen Chen, Bin Li 等ACM MM 2025
- From Pixels to Semantics: A Novel MLLM-Driven Approach for Explainable Tampered Text DetectionGuitao Xu, Ziqi Yi, Peirong Zhang, Jiahuan Cao 等ACM MM 2025 · 被引用 2 次
- D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image DewarpingHeng Li, Xiangping Wu, Qingcai ChenCVPR 2026
- Towards Generalized Physical Occlusion Detection On DocumentsYiang Zhu, Haoyue Wang, Zhenxing Qian, Sheng Li 等ACM MM 2025
