High-Resolution Document Shadow Removal via A Large-Scale Real-World Dataset and A Frequency-Aware Shadow Erasing Net
Zinuo Li, Xuhang Chen, Chi-Man Pun, Xiaodong Cun
摘要
Shadows often occur when we capture the document with casual equipment, which influences the visual quality and readability of the digital copies. Different from the algorithms for natural shadow removal, the algorithms in document shadow removal need to preserve the details of fonts and figures in high-resolution input. Previous works ignore this problem and remove the shadows via approximate attention and small datasets, which might not work in real-world situations. We handle high-resolution document shadow removal directly via a larger-scale real-world dataset and a carefully-designed frequency-aware network. As for the dataset, we acquire over 7k couples of high-resolution (2462 × 3699) images of real-world documents pairs with various samples under different lighting circumstances, which is 10 times larger than existing datasets. As for the design of the network, we decouple the high-resolution images in the frequency domain, where the low-frequency details and high-frequency boundaries can be effectively learned via the carefully designed network structure. Powered by our network and dataset, the proposed method shows a clearly better performance than previous methods in terms of visual quality and numerical results. The code, models, and dataset are available at https://github.com/CXH-Research/DocShadow-SD7K.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Uni-DocDiff: A Unified Document Restoration Model Based on DiffusionFangmin Zhao, Weichao Zeng, Zhenhang Li, Dongbao Yang 等ACM MM 2025 · 被引用 1 次
- DocRes: A Generalist Model Toward Unifying Document Image Restoration TasksJiaxin Zhang, Dezhi Peng, Chongyu Liu, Peirong Zhang 等CVPR 2024
- MMDIR: Multimodal Instruction-Driven Framework for Mixed-Degradation Document Image RestorationHeng Li, Xingyuan Wang, Yang Fan, Yunan Zhang 等CVPR 2026
它引用的顶会 Paper10
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou 等CVPR 2022 · 被引用 1,970 次
- Ultra-High-Definition Low-Light Image Enhancement: A Benchmark and Transformer-Based MethodTao Wang, Kaihao Zhang, Tianrun Shen, Wenhan Luo 等AAAI 2023 · 被引用 577 次
- LayoutLM: Pre-training of Text and Layout for Document Image UnderstandingYiheng Xu, Minghao Li, Lei Cui, Shaohan Huang 等KDD 2020 · 被引用 575 次
- Towards Ghost-Free Shadow Removal via Dual Hierarchical Aggregation Network and Shadow Matting GANXiaodong Cun, Chi-Man Pun, Cheng ShiAAAI 2020 · 被引用 272 次
相关 Paper
- BEDSR-Net: A Deep Shadow Removal Network From a Single Document ImageYun-Hsuan Lin, Wen-Chin Chen, Yung-Yu ChuangCVPR 2020
- OmniSR: Shadow Removal Under Direct and Indirect LightingJiamin Xu, Zelong Li, Yuxin Zheng, Chenyu Huang 等AAAI 2025 · 被引用 18 次
- Document Image Shadow Removal Guided by Color-Aware BackgroundLing Zhang, Yinghao He, Qing Zhang, Zheng Liu 等CVPR 2023
- FSR-Net: Deep Fourier Network for Shadow RemovalJun Yu, Peng He, Ziqi PengACM MM 2023 · 被引用 11 次
- Regional Attention For Shadow RemovalHengxing Liu, Mingjia Li, Xiaojie GuoACM MM 2024 · 被引用 10 次
