CRNet: A Center-aware Representation for Detecting Text of Arbitrary Shapes
Yu Zhou, Hongtao Xie, Shancheng Fang, Yan Li, Yongdong Zhang
Abstract
Existing scene text detection methods achieve state-of-the-art performance by designing elaborate anchors or complex post-processing. Nonetheless, most methods still face the dilemma of detecting adjacent texts as one instance and long text with large character spacing as multiple fragments. To tackle these problems, we propose an anchor-free scene text detector leveraging Center-aware Representation to achieve accurate arbitrary-shaped scene text detection namely CRNet. Firstly, we propose a center-aware location algorithm to explicitly learn center regions and center points of text instances, which is able to separate adjacent text instances effectively. Then, a multi-scale context extraction module capable of extracting local context, long-range dependencies and global context adaptively is designed to effectively perceive long text with large character spacing. Finally, a low-level features enhancement block is introduced to enhance the geometric information of text. Extensive experiments conducted on several benchmarks including SCUT-CTW1500, Total-Text, ICDAR2015, ICDAR2017 MLT, and MSRA-TD500 demonstrate the effectiveness of our method. Specifically, without any anchor and complicated post-processing, our CRNet achieves 84.2% and 85.1% on CTW1500 and MSRA-TD500 in F-measure, outperforming all state-of-the-art anchor-based and anchor-free methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers6
- TPSNet: Reverse Thinking of Thin Plate Splines for Arbitrary Shape Scene Text RepresentationWei Wang, Yu Zhou, Jiahao Lyu, Dayan Wu et al.ACM MM 2022 · 35 citations
- Mask is All You Need: Rethinking Mask R-CNN for Dense and Arbitrary-Shaped Scene Text DetectionXugong Qin, Yu Zhou, Youhui Guo, Dayan Wu et al.ACM MM 2021 · 33 citations
- Towards Robust Real-Time Scene Text Detection: From Semantic to Instance Representation LearningXugong Qin, Pengyuan Lyu, Chengquan Zhang, Yu Zhou et al.ACM MM 2023 · 21 citations
- LRANet: Towards Accurate and Efficient Scene Text Detection with Low-Rank Approximation NetworkYuchen Su, Zhineng Chen, Zhiwen Shao, Yuning Du et al.AAAI 2024 · 20 citations
- Robust Feature Rectification of Pretrained Vision Models for Object RecognitionShengchao Zhou, Gaofeng Meng, Zhaoxiang Zhang, Richard Yi Da Xu et al.AAAI 2023 · 1 citation
Related papers
- ContourNet: Taking a Further Step Toward Accurate Arbitrary-Shaped Scene Text DetectionYuxin Wang, Hongtao Xie, Zheng-Jun Zha, Mengting Xing et al.CVPR 2020
- Towards Unconstrained End-to-End Text SpottingSiyang Qin, Alessandro Bissacco, Michalis Raptis, Yasuhisa Fujii et al.ICCV 2019 · 138 citations
- CentripetalText: An Efficient Text Instance Representation for Scene Text DetectionTao Sheng, Jie Chen, Zhouhui LianNeurIPS 2021 · 30 citations
- Few Could Be Better Than All: Feature Sampling and Grouping for Scene Text DetectionJingqun Tang, Wenqing Zhang, Hongye Liu, Mingkun Yang et al.CVPR 2022 · 103 citations
- Progressive Contour Regression for Arbitrary-Shape Scene Text DetectionPengwen Dai, Sanyi Zhang, Hua Zhang, Xiaochun CaoCVPR 2021
