CMA: A Chromaticity Map Adapter for Robust Detection of Screen-Recapture Document Images
Changsheng Chen, Liangwei Lin, Yongqi Chen, Bin Li, Jishen Zeng, Jiwu Huang
Abstract
The rebroadcasting of screen-recaptured document images introduces a significant risk to the confidential documents processed in government departments and commercial companies. However, detecting recaptured document images subjected to distortions from online social networks (OSNs) is challenging since the common forensics cues, such as moiré pattern, are weakened during transmission. In this work, we first devise a pixel-level distortion model of the screen-recaptured document image to identify the robust features of color artifacts. Then, we extract a chromaticity map from the recaptured image to highlight the presence of color artifacts even under low-quality samples. Based on the prior understanding, we design a chromaticity map adapter (CMA) to efficiently extract the chromaticity map, and feed it into the transformer backbone as multi-modal prompt tokens. To evaluate the performance of the proposed method, we collect a recaptured office document image dataset with over 10K diverse samples. Experimental results demonstrate that the proposed CMA method outperforms a SOTA approach (with RGB modality only), reducing the average EER from 26.82% to 16.78%. Robustness evaluation shows that our method achieves 0.8688 and 0.7554 AUCs under samples with JPEG compression (QF=70) and resolution as low as 534×503 pixels.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fbf6d7db-11be-41bd-8894-fce616d4d1dfCited by top-tier papers1
- Detect Any AI-Counterfeited Text ImageChenfan Qu, Yiwu Zhong, Xuekang Zhu, Junchi Li et al.CVPR 2026
Builds on13
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Vision Transformer Adapter for Dense PredictionsZhe Chen, Yuchen Duan, Wenhai Wang, Junjun He et al.ICLR 2023 · 204 citations
- Z-Score Normalization, Hubness, and Few-Shot LearningNanyi Fei, Yizhao Gao, Zhiwu Lu, Tao XiangICCV 2021 · 158 citations
- Cascading Convolutional Color ConstancyHuanglin Yu, Ke Chen, Kaiqi Wang, Yanlin Qian et al.AAAI 2020 · 76 citations
- Rehearsal-Free Domain Continual Face Anti-Spoofing: Generalize More and Forget LessRizhao Cai, Yawen Cui, Zhi Li, Zitong Yu et al.ICCV 2023 · 29 citations
Related papers
- DSDNet: Raw Domain Demoiréing via Dual Color-Space SynergyQirui Yang, Fangpu Zhang, Yeying Jin, Qihua Cheng et al.ACM MM 2025 · 3 citations
- Unmask Tampering: Efficient Document Tampering Localization under Recapturing Attacks with Real Distortion KnowledgeChangsheng Chen, Wenyu Chen, Yinyin Lin, Bin Li et al.CCS 2025
- MoiréMop: Lightweight and Pervasive Moiré Removal for Mobile Camera-to-Screen InteractionZhaowei Wu, Jingyi Ning, Yanling Bu, Lei XieUbiComp 2026
- DocTr: Document Image Transformer for Geometric Unwarping and Illumination CorrectionHao Feng, Yuechen Wang, Wengang Zhou, Jiajun Deng et al.ACM MM 2021 · 66 citations
- Recaptured Raw Screen Image and Video Demoiréing via Channel and Spatial ModulationsYijia Cheng, Xin Liu, Jingyu YangNeurIPS 2023 · 23 citations
