Towards Ship License Plate Recognition in the Wild: A Large Benchmark and Strong Baseline
Baolong Liu, Ruiqing Yang, Roukai Huang, Wenhao Xu, Xin Pan, Chuanhuang Li, Bin Wang, Xun Wang, Jianfeng Dong
Abstract
The paper targets the challenging task of Ship License Plate (SLP) recognition. Existing methods for SLP recognition are hampered by the scarcity of large and publicly available datasets, leading to evaluations on small and non-representative datasets. To alleviate it, we have built a large dataset, called SLP34K, which consists of 34,385 images collected by an intelligent traffic surveillance system. The dataset is carefully manually annotated with text labels and attributes, and presents high data diversity by multiple installation locations and long capturing period of the cameras. Additionally, we propose a simple yet effective SLP recognition baseline method. The baseline is equipped with a strong visual encoder that benefits from initial pre-training via self-supervised learning, followed by further refinement through our devised semantic enhancement module. Extensive experiments on SLP34K verify the effectiveness of our proposed baseline. Moreover, while our baseline is designed for SLP recognition, it can also be used for common scene text recognition and achieve state-of-the-art performance on seven mainstream scene text recognition datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fdeaf1fd-415b-42e8-bd8e-87b07bc12e34Builds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- From Two to One: A New Scene Text Recognizer with Visual Language Modeling NetworkYuxin Wang, Hongtao Xie, Shancheng Fang, Jing Wang et al.ICCV 2021 · 184 citations
- Hierarchical Contrast for Unsupervised Skeleton-Based Action Representation LearningJianfeng Dong, Shengkai Sun, Zhonglin Liu, Shujie Chen et al.AAAI 2023 · 73 citations
- Revisiting Scene Text Recognition: A Data PerspectiveQing Jiang, Jiapeng Wang, Dezhi Peng, Chongyu Liu et al.ICCV 2023 · 70 citations
- Partially Relevant Video RetrievalJianfeng Dong, Xianke Chen, Minsong Zhang, Xun Yang et al.ACM MM 2022 · 65 citations
Related papers
- Advancing Ship Re-Identification in the Wild: The ShipReID-2400 Benchmark Dataset and D2InterNet Baseline MethodBaolong Liu, Roukai Huang, Xin Pan, Chuanhuang Li et al.SIGIR 2025 · 1 citation
- Vehicle Re-Identification in Aerial Imagery: Dataset and ApproachPeng Wang, Bingliang Jiao, Lu Yang, Yifei Yang et al.ICCV 2019 · 67 citations
- Traffic Scene Parsing Through the TSP6K DatasetPeng-Tao Jiang, Yuqi Yang, Yang Cao, Qibin Hou et al.CVPR 2024
- Structural Information Guided Multimodal Pre-training for Vehicle-Centric PerceptionXiao Wang, Wentao Wu, Chenglong Li, Zhicheng Zhao et al.AAAI 2024 · 10 citations
- LP-Diff: Towards Improved Restoration of Real-World Degraded License PlateHaoyan Gong, Zhenrong Zhang, Yuzheng Feng, Anh Nguyen et al.CVPR 2025
