High-Resolution Document Shadow Removal via A Large-Scale Real-World Dataset and A Frequency-Aware Shadow Erasing Net
Zinuo Li, Xuhang Chen, Chi-Man Pun, Xiaodong Cun
Abstract
Shadows often occur when we capture the document with casual equipment, which influences the visual quality and readability of the digital copies. Different from the algorithms for natural shadow removal, the algorithms in document shadow removal need to preserve the details of fonts and figures in high-resolution input. Previous works ignore this problem and remove the shadows via approximate attention and small datasets, which might not work in real-world situations. We handle high-resolution document shadow removal directly via a larger-scale real-world dataset and a carefully-designed frequency-aware network. As for the dataset, we acquire over 7k couples of high-resolution (2462 × 3699) images of real-world documents pairs with various samples under different lighting circumstances, which is 10 times larger than existing datasets. As for the design of the network, we decouple the high-resolution images in the frequency domain, where the low-frequency details and high-frequency boundaries can be effectively learned via the carefully designed network structure. Powered by our network and dataset, the proposed method shows a clearly better performance than previous methods in terms of visual quality and numerical results. The code, models, and dataset are available at https://github.com/CXH-Research/DocShadow-SD7K.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30b5022f-90ba-41a0-b762-7d6f7d81ed3dCited by top-tier papers3
- Uni-DocDiff: A Unified Document Restoration Model Based on DiffusionFangmin Zhao, Weichao Zeng, Zhenhang Li, Dongbao Yang et al.ACM MM 2025 · 1 citation
- DocRes: A Generalist Model Toward Unifying Document Image Restoration TasksJiaxin Zhang, Dezhi Peng, Chongyu Liu, Peirong Zhang et al.CVPR 2024
- MMDIR: Multimodal Instruction-Driven Framework for Mixed-Degradation Document Image RestorationHeng Li, Xingyuan Wang, Yang Fan, Yunan Zhang et al.CVPR 2026
Builds on10
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- Ultra-High-Definition Low-Light Image Enhancement: A Benchmark and Transformer-Based MethodTao Wang, Kaihao Zhang, Tianrun Shen, Wenhan Luo et al.AAAI 2023 · 577 citations
- LayoutLM: Pre-training of Text and Layout for Document Image UnderstandingYiheng Xu, Minghao Li, Lei Cui, Shaohan Huang et al.KDD 2020 · 575 citations
- Towards Ghost-Free Shadow Removal via Dual Hierarchical Aggregation Network and Shadow Matting GANXiaodong Cun, Chi-Man Pun, Cheng ShiAAAI 2020 · 272 citations
Related papers
- BEDSR-Net: A Deep Shadow Removal Network From a Single Document ImageYun-Hsuan Lin, Wen-Chin Chen, Yung-Yu ChuangCVPR 2020
- OmniSR: Shadow Removal Under Direct and Indirect LightingJiamin Xu, Zelong Li, Yuxin Zheng, Chenyu Huang et al.AAAI 2025 · 18 citations
- Document Image Shadow Removal Guided by Color-Aware BackgroundLing Zhang, Yinghao He, Qing Zhang, Zheng Liu et al.CVPR 2023
- FSR-Net: Deep Fourier Network for Shadow RemovalJun Yu, Peng He, Ziqi PengACM MM 2023 · 11 citations
- Regional Attention For Shadow RemovalHengxing Liu, Mingjia Li, Xiaojie GuoACM MM 2024 · 10 citations
