Portable Active Learning for Object Detection
Rashi Sharma, Justin Timothy C. Bersamin, Karthikk Subramanian
Abstract
Annotating bounding boxes is costly and limits the scalability of object detection. This challenge is compounded by the need to preserve high accuracy while minimizing manual effort in real-world applications. Prior active learning methods often depend on model features or modify detector internals and training schedules, increasing integration overhead. Moreover, they rarely jointly exploit the benefits of image-level signals, class-imbalance cues, and instancelevel uncertainty for comprehensive selection. We present Portable Active Learning (PAL), a detector-agnostic, easily portable framework that operates solely on inference outputs. PAL combines class-wise instance uncertainty with image-level diversity to guide data selection. At each round, PAL trains lightweight class-specific logistic classifiers to distinguish true from false positives, producing entropybased uncertainty scores for proposals. Candidate images are then refined using global image entropy, class diversity, and image similarity, yielding batches that are both informative and diverse. PAL requires no changes to model internals or training pipelines, ensuring broad compatibility across detectors. Extensive experiments on COCO, PASCAL VOC, and BDD100K demonstrate that PAL consistently improves label efficiency and detection accuracy compared to existing active learning baselines, making it a practical solution for scalable and cost-effective deployment of object detection in real-world settings.
- Work done while an intern at Panasonic R&D Center Singapore. See author contributions for further details.
often becoming a bottleneck in deploying object detection systems for new domains or rare categories.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on9
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Active Learning for Deep Object Detection via Probabilistic ModelingJiwoong Choi, Ismail Elezi, Hyuk-Jae Lee, Clément Farabet et al.ICCV 2021 · 144 citations
- Entropy-based Active Learning for Object Detection with Progressive Diversity ConstraintJiaxi Wu, Jiaxin Chen, Di HuangCVPR 2022 · 88 citations
- Active Teacher for Semi-Supervised Object DetectionPeng Mi, Jianghang Lin, Yiyi Zhou, Yunhang Shen et al.CVPR 2022 · 83 citations
Related papers
- Plug and Play Active Learning for Object DetectionChenhongyi Yang, Lichao Huang, Elliot J. CrowleyCVPR 2024 · 29 citations
- Not All Labels Are Equal: Rationalizing The Labeling Costs for Training Object DetectionIsmail Elezi, Zhiding Yu, Anima Anandkumar, Laura Leal-Taixé et al.CVPR 2022 · 45 citations
- Not All Out-of-Distribution Data Are Harmful to Open-Set Active LearningYang Yang, Yuxuan Zhang, Xin Song, Yi XuNeurIPS 2023 · 48 citations
- Box-Level Active DetectionMengyao Lyu, Jundong Zhou, Hui Chen, Yijie Huang et al.CVPR 2023
- Multiple Instance Active Learning for Object DetectionTianning Yuan, Fang Wan, Mengying Fu, Jianzhuang Liu et al.CVPR 2021
