End-to-end Piece-wise Unwarping of Document Images
Sagnik Das, Kunwar Yashraj Singh, Jon Wu, Erhan Bas, Vijay Mahadevan, Rahul Bhotika, Dimitris Samaras
摘要
Document unwarping attempts to undo physical deformations of the paper and recover a 'flatbed' scanned document-image for downstream tasks such as OCR. Current state-of-the-art relies on global unwarping of the document which is not robust to local deformation changes. Moreover, a global unwarping often produces spurious warping artifacts in less warped regions to compensate for severe warps present in other parts of the document. In this paper, we propose the first end-to-end trainable piece-wise unwarping 1 method that predicts local deformation fields and stitches them together with global information to obtain an improved unwarping. The proposed piece-wise formulation results in 4% improvement in terms of multi-scale structural similarity (MS-SSIM) and shows better performance in terms of OCR metrics, character error rate (CER) and word error rate (WER) compared to the state-of-the-art.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Revisiting Document Image Dewarping by Grid RegularizationXiangwei Jiang, Rujiao Long, Nan Xue, Zhibo Yang 等CVPR 2022 · 被引用 39 次
- Learning From Documents in the Wild to Improve Document UnwarpingKe Ma, Sagnik Das, Zhixin Shu, Dimitris SamarasSIGGRAPH 2022 · 被引用 39 次
- Foreground and Text-lines Aware Document Image RectificationHeng Li, Xiangping Wu, Qingcai Chen, Qianjin XiangICCV 2023 · 被引用 21 次
- ForCenNet: Foreground-Centric Network for Document Image RectificationPeng Cai, Qiang Li, Kaicheng Yang, Dong Guo 等ICCV 2025 · 被引用 1 次
- D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image DewarpingHeng Li, Xiangping Wu, Qingcai ChenCVPR 2026
它引用的顶会 Paper1
相关 Paper
- DocTr: Document Image Transformer for Geometric Unwarping and Illumination CorrectionHao Feng, Yuechen Wang, Wengang Zhou, Jiajun Deng 等ACM MM 2021 · 被引用 66 次
- Document Registration: Towards Automated Labeling of Pixel-Level Alignment Between Warped-Flat DocumentsWeiguang Zhang, Qiufeng Wang, Kaizhu Huang, Xiaowei Huang 等ACM MM 2024 · 被引用 1 次
- TrOCR: Transformer-Based Optical Character Recognition with Pre-trained ModelsMinghao Li, Tengchao Lv, Jingye Chen, Lei Cui 等AAAI 2023 · 被引用 607 次
- Towards Unconstrained End-to-End Text SpottingSiyang Qin, Alessandro Bissacco, Michalis Raptis, Yasuhisa Fujii 等ICCV 2019 · 被引用 138 次
- Axis-Aligned Document DewarpingChaoyun Wang, I-Chao Shen, Takeo Igarashi, Caigui JiangAAAI 2026
