Uncertainty Meets Diversity: A Comprehensive Active Learning Framework for Indoor 3D Object Detection
Jiangyi Wang, Na Zhao
Abstract
Active learning has emerged as a promising approach to reduce the substantial annotation burden in 3D object detection tasks, spurring several initiatives in outdoor environments. However, its application in indoor environments remains unexplored. Compared to outdoor 3D datasets, indoor datasets face significant challenges, including fewer training samples per class, a greater number of classes, more severe class imbalance, and more diverse scene types and intra-class variances. This paper presents the first study on active learning for indoor 3D object detection, where we propose a novel framework tailored for this task. Our method incorporates two key criteria -uncertainty and diversity -to actively select the most ambiguous and informative unlabeled samples for annotation. The uncertainty criterion accounts for both inaccurate detections and undetected objects, ensuring that the most ambiguous samples are prioritized. Meanwhile, the diversity criterion is formulated as a joint optimization problem that maximizes the diversity of both object class distributions and scene types, using a new Class-aware Adaptive Prototype (CAP) bank. The CAP bank dynamically allocates representative prototypes to each class, helping to capture varying intra-class diversity across different categories. We evaluate our method on SUN RGB-D and ScanNetV2, where it outperforms baselines by a significant margin, achieving over 85% of fully-supervised performance with just 10% of the annotation budget. Our code is available at: https://github.com/JoeWang-0519/CVPR25 UMD.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3e1d65f4-6ffc-436c-a62c-7c3b1d79df73Cited by top-tier papers5
- AffordBot: 3D Fine-grained Embodied Reasoning via Multimodal Large Language ModelsXinyi Wang, Xun Yang, Yanlong Xu, Yuchen Wu et al.NeurIPS 2025 · 18 citations
- CCF: Complementary Collaborative Fusion for Domain Generalized Multi-Modal 3D Object DetectionYuchen Wu, Kun Wang, Yining Pan, Na ZhaoCVPR 2026 · 4 citations
- Few-Shot Incremental 3D Object Detection in Dynamic Indoor EnvironmentsYun Zhu, Jianjun Qian, Jian Yang, Jin Xie et al.CVPR 2026 · 2 citations
- Graph Smoothing for Enhanced Local Geometry Learning in Point Cloud AnalysisShangbo Yuan, Jie Xu, Ping Hu, Xiaofeng Zhu et al.AAAI 2026
- FlyMeThrough: Human-AI Collaborative 3D Indoor Mapping with Commodity DronesXia Su, Ruiqi Chen, Jingwei Ma, Chu Li et al.UIST 2025
Builds on27
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 662 citations
- An End-to-End Transformer Model for 3D Object DetectionIshan Misra, Rohit Girdhar, Armand JoulinICCV 2021 · 602 citations
- Batch Active Learning at ScaleGui Citovsky, Giulia DeSalvo, Claudio Gentile, Lazaros Karydas et al.NeurIPS 2021 · 220 citations
Related papers
- Entropy-based Active Learning for Object Detection with Progressive Diversity ConstraintJiaxi Wu, Jiaxin Chen, Di HuangCVPR 2022 · 88 citations
- Learning Class Prototypes for Unified Sparse-Supervised 3D Object DetectionYun Zhu, Le Hui, Hang Yang, Jianjun Qian et al.CVPR 2025
- Exploring Active 3D Object Detection from a Generalization PerspectiveYadan Luo, Zhuoxiao Chen, Zijian Wang, Xin Yu et al.ICLR 2023 · 3 citations
- ReDAL: Region-based and Diversity-aware Active Learning for Point Cloud Semantic SegmentationTsung-Han Wu, Yueh-Cheng Liu, Yu-Kai Huang, Hsin-Ying Lee et al.ICCV 2021 · 92 citations
- Plug and Play Active Learning for Object DetectionChenhongyi Yang, Lichao Huang, Elliot J. CrowleyCVPR 2024 · 29 citations
