DITL2: Dual-Stage Invariance Transfer Learning for Generalizable Document Image Tampering Localization
Songze Li, Yunfei Guo, Shen Chen, Bin Li, Kaiqing Lin, Changsheng Chen, Haodong Li, Taiping Yao, Shouhong Ding
摘要
Document Image Tampering Localization (DITL) advances considerably, yet achieving robust cross-dataset generalization remains a formidable challenge for practical applications. Expanding existing document datasets for training is labor-intensive, making it appealing to incorporate data from non-document domains such as natural scene images. However, domain-specific variations, including differences in color distribution and texture, compromise the performance of joint training. To address this issue, we propose DITL2, a Dual-stage Invariance Transfer Learning framework for Document Image Tampering Localization that consists of Cross-Domain Invariance Pre-training (CDIP) and Frequency Decoupling Parameter Adaptation (FDPA). In the pre-training stage, CDIP employs style transfer and texture consistency learning to suppress domain-specific influences from tampered natural scene images, and tampering trace commonality learning to acquire domain-invariant features. In the fine-tuning stage, FDPA adapts the parameters of the pre-trained model, leveraging the general knowledge from the pre-trained model to address DITL tasks while reducing the risk of overfitting. Experiments show that this approach effectively leverages external data resources to boost model performance, achieving state-of-the-art results across a variety of cross-dataset settings.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- Deep Residual Injection for Full-Spectrum Forensic Signal Perception in Multimodal Large Language ModelsKaiqing Lin, Zhiyuan Yan, Ruoxin Chen, Ke-Yue Zhang 等ICML 2026 · 被引用 1 次
- Detect Any AI-Counterfeited Text ImageChenfan Qu, Yiwu Zhong, Xuekang Zhu, Junchi Li 等CVPR 2026
相关 Paper
- Towards Robust Tampered Text Detection in Document Image: New Dataset and New SolutionChenfan Qu, Chongyu Liu, Yuliang Liu, Xinhong Chen 等CVPR 2023
- ADCD-Net: Robust Document Image Forgery Localization via Adaptive DCT Feature and Hierarchical Content DisentanglementKahim Wong, Jicheng Zhou, Haiwei Wu, Yain-Whar Si 等ICCV 2025 · 被引用 3 次
- Towards Better Robustness Against Natural Corruptions in Document Tampering LocalizationHuiru Shao, Kaizhu Huang, Wei Wang, Xiaowei Huang 等AAAI 2025 · 被引用 1 次
- LDP: Generalizing to Multilingual Visual Information Extraction by Language Decoupled PretrainingHuawen Shen, Gengluo Li, Jinwen Zhong, Yu ZhouAAAI 2025 · 被引用 4 次
- DSD-DA: Distillation-based Source Debiasing for Domain Adaptive Object DetectionYongchao Feng, Shiwei Li, Yingjie Gao, Ziyue Huang 等ICML 2024 · 被引用 11 次
