Logo-2K+: A Large-Scale Logo Dataset for Scalable Logo Classification
Jing Wang, Weiqing Min, Sujuan Hou, Shengnan Ma, Yuanjie Zheng, Haishuai Wang, Shuqiang Jiang
摘要
Logo classification has gained increasing attention for its various applications, such as copyright infringement detection, product recommendation and contextual advertising. Compared with other types of object images, the real-world logo images have larger variety in logo appearance and more complexity in their background. Therefore, recognizing the logo from images is challenging. To support efforts towards scalable logo classification task, we have curated a dataset, Logo-2K+, a new large-scale publicly available real-world logo dataset with 2,341 categories and 167,140 images. Compared with existing popular logo datasets, such as FlickrLogos-32 and LOGO-Net, Logo-2K+ has more comprehensive coverage of logo categories and larger quantity of logo images. Moreover, we propose a Discriminative Region Navigation and Augmentation Network (DRNA-Net), which is capable of discovering more informative logo regions and augmenting these image regions for logo classification. DRNA-Net consists of four sub-networks: the navigator sub-network first selected informative logo-relevant regions guided by the teacher sub-network, which can evaluate its confidence belonging to the ground-truth logo class. The data augmentation sub-network then augments the selected regions via both region cropping and region dropping. Finally, the scrutinizer sub-network fuses features from augmented regions and the whole image for logo classification. Comprehensive experiments on Logo-2K+ and other three existing benchmark datasets demonstrate the effectiveness of proposed method. Logo-2K+ and the proposed strong baseline DRNA-Net are expected to further the development of scalable logo image recognition, and the Logo-2K+ dataset can be found at https://github.com/msn199959/Logo-2k-plus-Dataset .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Phishpedia: A Hybrid Deep Learning Based Approach to Visually Identify Phishing WebpagesYun Lin, Ruofan Liu, Dinil Mon Divakaran, Jun Yang Ng 等USENIX Security 2021 · 被引用 164 次
- FoodLogoDet-1500: A Dataset for Large-Scale Food Logo Detection via Multi-Scale Feature Decoupling NetworkQiang Hou, Weiqing Min, Jing Wang, Sujuan Hou 等ACM MM 2021 · 被引用 28 次
- Meta Learning on a Sequence of Imbalanced Domains with Difficulty AwarenessZhenyi Wang, Tiehang Duan, Le Fang, Qiuling Suo 等ICCV 2021 · 被引用 21 次
- Learning to Learn and Remember Super Long Multi-Domain Task SequenceZhenyi Wang, Li Shen, Tiehang Duan, Donglin Zhan 等CVPR 2022 · 被引用 19 次
- Safe-SD: Safe and Traceable Stable Diffusion with Text Prompt Trigger for Invisible Generative WatermarkingZhiyuan Ma, Guoli Jia, Biqing Qi, Bowen ZhouACM MM 2024 · 被引用 14 次
相关 Paper
- Cross-View Representation Learning for Multi-View Logo Classification with Information BottleneckJing Wang, Yuanjie Zheng, Jingqi Song, Sujuan HouACM MM 2021 · 被引用 8 次
- Aesthetic Text Logo Synthesis via Content-aware Layout InferringYizhi Wang, Guo Pu, Wenhan Luo, Yexin Wang 等CVPR 2022 · 被引用 27 次
- GLDesigner: Leveraging Multi-Modal LLMs as Designer for Enhanced Aesthetic Text Glyph LayoutsJunwen He, Yifan Wang, Lijun Wang, Huchuan Lu 等ACM MM 2025 · 被引用 1 次
- Cycle-Consistent Tuning for Layered Image DecompositionZheng Gu, Min Lu, Zhida Sun, Dani Lischinski 等CVPR 2026
- Leverage Your Local and Global Representations: A New Self-Supervised Learning StrategyTong Zhang, Congpei Qiu, Wei Ke, Sabine Süsstrunk 等CVPR 2022 · 被引用 24 次
