DLWL: Improving Detection for Lowshot Classes With Weakly Labelled Data
Vignesh Ramanathan, Rui Wang, Dhruv Mahajan
Abstract
Large detection datasets have a long tail of lowshot classes with very few bounding box annotations. We wish to improve detection for lowshot classes with weakly labelled web-scale datasets only having image-level labels. This requires a detection framework that can be jointly trained with limited number of bounding box annotated images and large number of weakly labelled images. Towards this end, we propose a modification to the FRCNN [39] model to automatically infer label assignment for objects proposals from weakly labelled images during training. We pose this label assignment as a Linear Program with constraints on the number and overlap of object instances in an image. We show that this can be solved efficiently during training for weakly labelled images. Compared to just training with few annotated examples, augmenting with weakly labelled examples in our framework provides significant gains. We demonstrate this on the LVIS dataset (3.5% gain in AP) as well as different lowshot variants of the COCO dataset. We provide a thorough analysis of the effect of amount of weakly labelled and fully labelled data required to train the detection model. Our DLWL framework can also outperform self-supervised baselines like omni-supervision [37] for lowshot classes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c7a7f661-bc89-4cae-8a89-5be5a081e035Cited by top-tier papers10
- Scaling Open-Vocabulary Object DetectionMatthias Minderer, Alexey A. Gritsenko, Neil HoulsbyNeurIPS 2023 · 482 citations
- Bridging the Gap between Object and Image-level Representations for Open-Vocabulary DetectionHanoona Abdul Rasheed, Muhammad Maaz, Muhammad Uzair Khattak, Salman H. Khan et al.NeurIPS 2022 · 215 citations
- On Model Calibration for Long-Tailed Object Detection and Instance SegmentationTai-Yu Pan, Cheng Zhang, Yandong Li, Hexiang Hu et al.NeurIPS 2021 · 56 citations
- MosaicOS: A Simple and Effective Use of Object-Centric Images for Long-Tailed Object DetectionCheng Zhang, Tai-Yu Pan, Yandong Li, Hexiang Hu et al.ICCV 2021 · 50 citations
- Betrayed by Captions: Joint Caption Grounding and Generation for Open Vocabulary Instance SegmentationJianzong Wu, Xiangtai Li, Henghui Ding, Xia Li et al.ICCV 2023 · 36 citations
Builds on9
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu et al.ICCV 2019 · 835 citations
- Meta-Learning to Detect Rare ObjectsYu-Xiong Wang, Deva Ramanan, Martial HebertICCV 2019 · 339 citations
- WSOD2: Learning Bottom-Up and Top-Down Objectness Distillation for Weakly-Supervised Object DetectionZhaoyang Zeng, Bei Liu, Jianlong Fu, Hongyang Chao et al.ICCV 2019 · 162 citations
- Towards Precise End-to-End Weakly Supervised Object Detection NetworkKe Yang, Dongsheng Li, Yong DouICCV 2019 · 141 citations
- Weakly Supervised Object Detection With Segmentation CollaborationXiaoyan Li, Meina Kan, Shiguang Shan, Xilin ChenICCV 2019 · 105 citations
Related papers
- UniT: Unified Knowledge Transfer for Any-Shot Object Detection and SegmentationSiddhesh Khandelwal, Raghav Goyal, Leonid SigalCVPR 2021
- SimLTD: Simple Supervised and Semi-Supervised Long-Tailed Object DetectionPhi Vu TranCVPR 2025
- Detecting 11K Classes: Large Scale Object Detection Without Fine-Grained Bounding BoxesHao Yang, Hao Wu, Hao ChenICCV 2019 · 30 citations
- iFS-RCNN: An Incremental Few-shot Instance SegmenterKhoi Nguyen, Sinisa TodorovicCVPR 2022 · 23 citations
- Open-Vocabulary Object Detection via Language HierarchyJiaxing Huang, Jingyi Zhang, Kai Jiang, Shijian LuNeurIPS 2024 · 16 citations
