ForCenNet: Foreground-Centric Network for Document Image Rectification
Peng Cai, Qiang Li, Kaicheng Yang, Dong Guo, Jia Li, Nan Zhou, Xiang An, Ninghua Yang, Jiankang Deng
摘要
Document image rectification aims to eliminate geometric deformation in photographed documents to facilitate text recognition. However, existing methods often neglect the significance of foreground elements, which provide essential geometric references and layout information for document image correction. In this paper, we introduce Foreground-Centric Network (ForCenNet) to eliminate geometric distortions in document images. Specifically, we initially propose a foreground-centric label generation method, which extracts detailed foreground elements from an undistorted image. Then we introduce a foreground-centric mask mechanism to enhance the distinction between readable and background regions. Furthermore, we design a curvature consistency loss to leverage the detailed foreground labels to help the model understand the distorted geometric distribution. Extensive experiments demonstrate that ForCenNet achieves new state-of-the-art on four real-world benchmarks, such as DocUNet, DIR300, WarpDoc, and DocReal. Quantitative analysis shows that the proposed method effectively undistorts layout elements, such as text lines and table borders. The resources for further comparison are provided at https://github.com/caipeng328/ ForCenNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- DewarpNet: Single-Image Document Unwarping With Stacked 3D and 2D Regression NetworksSagnik Das, Ke Ma, Zhixin Shu, Dimitris Samaras 等ICCV 2019 · 被引用 97 次
- DocTr: Document Image Transformer for Geometric Unwarping and Illumination CorrectionHao Feng, Yuechen Wang, Wengang Zhou, Jiajun Deng 等ACM MM 2021 · 被引用 66 次
- End-to-end Piece-wise Unwarping of Document ImagesSagnik Das, Kunwar Yashraj Singh, Jon Wu, Erhan Bas 等ICCV 2021 · 被引用 41 次
相关 Paper
- Foreground and Text-lines Aware Document Image RectificationHeng Li, Xiangping Wu, Qingcai Chen, Qianjin XiangICCV 2023 · 被引用 21 次
- Revisiting Document Image Dewarping by Grid RegularizationXiangwei Jiang, Rujiao Long, Nan Xue, Zhibo Yang 等CVPR 2022 · 被引用 39 次
- Fourier Document Restoration for Robust Document Dewarping and RecognitionChuhui Xue, Zichen Tian, Fangneng Zhan, Shijian Lu 等CVPR 2022 · 被引用 37 次
- Document Registration: Towards Automated Labeling of Pixel-Level Alignment Between Warped-Flat DocumentsWeiguang Zhang, Qiufeng Wang, Kaizhu Huang, Xiaowei Huang 等ACM MM 2024 · 被引用 1 次
- Symmetry-Constrained Rectification Network for Scene Text RecognitionMingkun Yang, Yushuo Guan, Minghui Liao, Xin He 等ICCV 2019 · 被引用 136 次
