MOST: A Multi-Oriented Scene Text Detector With Localization Refinement
Minghang He, Minghui Liao, Zhibo Yang, Humen Zhong, Jun Tang, Wenqing Cheng, Cong Yao, Yongpan Wang, Xiang Bai
摘要
Over the past few years, the field of scene text detection has progressed rapidly that modern text detectors are able to hunt text in various challenging scenarios. However, they might still fall short when handling text instances of extreme aspect ratios and varying scales. To tackle such difficulties, we propose in this paper a new algorithm for scene text detection, which puts forward a set of strategies to significantly improve the quality of text localization. Specifically, a Text Feature Alignment Module (TFAM) is proposed to dynamically adjust the receptive fields of features based on initial raw detections; a Position-Aware Non-Maximum Suppression (PA-NMS) module is devised to selectively concentrate on reliable raw detections and exclude unreliable ones; besides, we propose an Instancewise IoU loss for balanced training to deal with text instances of different scales. An extensive ablation study demonstrates the effectiveness and superiority of the proposed strategies. The resulting text detection system, which integrates the proposed strategies with a leading scene text detector EAST, achieves state-of-the-art or competitive performance on various standard benchmarks for text detection while keeping a fast running speed.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Few Could Be Better Than All: Feature Sampling and Grouping for Scene Text DetectionJingqun Tang, Wenqing Zhang, Hongye Liu, Mingkun Yang 等CVPR 2022 · 被引用 103 次
- Towards End-to-End Unified Scene Text Detection and Layout AnalysisShangbang Long, Siyang Qin, Dmitry Panteleev, Alessandro Bissacco 等CVPR 2022 · 被引用 86 次
- ESTextSpotter: Towards Better Scene Text Spotting with Explicit Synergy in TransformerMingxin Huang, Jiaxin Zhang, Dezhi Peng, Hao Lu 等ICCV 2023 · 被引用 44 次
- Vision-Language Pre-Training for Boosting Scene Text DetectorsSibo Song, Jianqiang Wan, Zhibo Yang, Jun Tang 等CVPR 2022 · 被引用 38 次
- TPSNet: Reverse Thinking of Thin Plate Splines for Arbitrary Shape Scene Text RepresentationWei Wang, Yu Zhou, Jiahao Lyu, Dayan Wu 等ACM MM 2022 · 被引用 35 次
它引用的顶会 Paper5
- Real-Time Scene Text Detection with Differentiable BinarizationMinghui Liao, Zhaoyi Wan, Cong Yao, Kai Chen 等AAAI 2020 · 被引用 818 次
- Efficient and Accurate Arbitrary-Shaped Text Detection With Pixel Aggregation NetworkWenhai Wang, Enze Xie, Xiaoge Song, Yuhang Zang 等ICCV 2019 · 被引用 490 次
- Geometry Normalization Networks for Accurate Scene Text DetectionJiaqi Duan, Youjiang Xu, Zhanghui Kuang, Xiaoyu Yue 等ICCV 2019 · 被引用 33 次
- Deep Relational Reasoning Graph Network for Arbitrary Shape Text DetectionShi-Xue Zhang, Xiaobin Zhu, Jie-Bo Hou, Chang Liu 等CVPR 2020
- ContourNet: Taking a Further Step Toward Accurate Arbitrary-Shaped Scene Text DetectionYuxin Wang, Hongtao Xie, Zheng-Jun Zha, Mengting Xing 等CVPR 2020
相关 Paper
- CRNet: A Center-aware Representation for Detecting Text of Arbitrary ShapesYu Zhou, Hongtao Xie, Shancheng Fang, Yan Li 等ACM MM 2020 · 被引用 31 次
- Kernel Adaptive Convolution for Scene Text Detection via Distance Map PredictionJinzhi Zheng, Heng Fan, Libo ZhangCVPR 2024 · 被引用 14 次
- TDI TextSpotter: Taking Data Imbalance into Account in Scene Text SpottingYu Zhou, Hongtao Xie, Shancheng Fang, Jing Wang 等ACM MM 2021 · 被引用 4 次
- TextScanner: Reading Characters in Order for Robust Scene Text RecognitionZhaoyi Wan, Minghang He, Haoran Chen, Xiang Bai 等AAAI 2020 · 被引用 158 次
- Perceiving Ambiguity and Semantics without Recognition: An Efficient and Effective Ambiguous Scene Text DetectorYan Shu, Wei Wang, Yu Zhou, Shaohui Liu 等ACM MM 2023 · 被引用 9 次
