Dataset Preparation for Arbitrary Object Detection: An Automatic Approach based on Web Information in English
Shucheng Li, Boyu Chang, Bo Yang, Hao Wu, Sheng Zhong, Fengyuan Xu
摘要
Automatic dataset preparation can help users avoid labor-intensive and costly manual data annotations. The difficulty in preparing a high-quality dataset for object detection involves three key aspects: relevance, naturality, and balance, which are not addressed by existing works. In this paper, we leverage information from the web, and propose a fully-automatic dataset preparation mechanism without any human annotation, which can automatically prepare a high-quality training dataset for the detection task with English text terms describing target objects. It contains three key designs, i.e., keyword expansion, data de-noising, and data balancing. Our experiments demonstrate that the object detectors trained with auto-prepared data are comparable to those trained with benchmark datasets and outperform other baselines. We also demonstrate the effectiveness of our approach in several more challenging real-world object categories that are not included in the benchmark datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- Comprehensive Attention Self-Distillation for Weakly-Supervised Object DetectionZeyi Huang, Yang Zou, B. V. K. Vijaya Kumar, Dong HuangNeurIPS 2020 · 被引用 149 次
- C-MIDN: Coupled Multiple Instance Detection Network With Segmentation Guidance for Weakly Supervised Object DetectionGao Yan, Boxiao Liu, Nan Guo, Xiaochun Ye 等ICCV 2019 · 被引用 130 次
- Instance Mining with Class Feature Banks for Weakly Supervised Object DetectionYufei Yin, Jiajun Deng, Wengang Zhou, Houqiang LiAAAI 2021 · 被引用 46 次
相关 Paper
- Simple Multi-dataset DetectionXingyi Zhou, Vladlen Koltun, Philipp KrähenbühlCVPR 2022 · 被引用 85 次
- ImaginaryNet: Learning Object Detectors without Real Images and AnnotationsMinheng Ni, Zitong Huang, Kailai Feng, Wangmeng ZuoICLR 2023 · 被引用 5 次
- nocaps: novel object captioning at scaleHarsh Agrawal, Peter Anderson, Karan Desai, Yufei Wang 等ICCV 2019 · 被引用 631 次
- Cap2Det: Learning to Amplify Weak Caption Supervision for Object DetectionKeren Ye, Mingda Zhang, Adriana Kovashka, Wei Li 等ICCV 2019 · 被引用 61 次
- OmniLabel: A Challenging Benchmark for Language-Based Object DetectionSamuel Schulter, Vijay Kumar B. G, Yumin Suh, Konstantinos M. Dafnis 等ICCV 2023 · 被引用 18 次
