Harnessing Uncertainty-Aware Bounding Boxes for Unsupervised 3D Object Detection
Ruiyang Zhang, Hu Zhang, Zhedong Zheng
Abstract
Unsupervised 3D object detection aims to identify objects of interest from unlabeled raw data, such as LiDAR points. Recent approaches usually adopt pseudo 3D bounding boxes (3D bboxes) from clustering algorithm to initialize the model training. However, pseudo bboxes inevitably contain noise, and such inaccuracies accumulate to the final model, compromising the performance. Therefore, in an attempt to mitigate the negative impact of inaccurate pseudo bboxes, we introduce a new uncertainty-aware framework for unsupervised 3D object detection, dubbed UA3D. In particular, our method consists of two phases: uncertainty estimation and uncertainty regularization. (1) In the uncertainty estimation phase, we incorporate an extra auxiliary detection branch alongside the original primary detector. The prediction disparity between the primary and auxiliary detectors could reflect fine-grained uncertainty at the box coordinate level. (2) Based on the assessed uncertainty, we adaptively adjust the weight of every 3D bbox coordinate via uncertainty regularization, refining the training process on pseudo bboxes. For pseudo bbox coordinate with high uncertainty, we assign a relatively low loss weight. Extensive experiments verify that the proposed method is robust against the noisy pseudo bboxes, yielding substantial improvements on nuScenes and Lyft compared to existing approaches, with increases of +6.9% AP and +2.5% AP on nuScenes, and +4.1% AP and +2.0% AP on Lyft.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7888d759-80b4-4f0f-9bcc-9f8ff7603b3fCited by top-tier papers2
- Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward ModelingGuiyu Zhang, Huan-ang Gao, Zijian Jiang, Hao Zhao et al.ICLR 2025
- SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal ModelsRuiyang Zhang, Dongzhan Zhou, Zhedong ZhengICLR 2026
Builds on18
- Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMsMiao Xiong, Zhiyuan Hu, Xinyang Lu, Yifei Li et al.ICLR 2024 · 867 citations
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen et al.ICCV 2019 · 840 citations
- Gaussian YOLOv3: An Accurate and Fast Object Detector Using Localization Uncertainty for Autonomous DrivingJiwoong Choi, Dayoung Chun, Hyun Kim, Hyuk-Jae LeeICCV 2019 · 445 citations
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski et al.ICML 2020 · 409 citations
- Ensemble Distribution DistillationAndrey Malinin, Bruno Mlodozeniec, Mark J. F. GalesICLR 2020 · 273 citations
Related papers
- Not Every Side Is Equal: Localization Uncertainty Estimation for Semi-Supervised 3D Object DetectionChuxin Wang, Wenfei Yang, Tianzhu ZhangICCV 2023 · 9 citations
- AnnofreeOD: Detecting All Classes at Low Frame Rates Without Human AnnotationsBoyi Sun, Yuhang Liu, Houxin He, Yonglin Tian et al.ICCV 2025 · 1 citation
- OpenBox: Annotate Any Bounding Boxes in 3DIn-Jae Lee, Mungyeom Kim, Kwonyoung Ryu, Pierre Musacchio et al.NeurIPS 2025 · 7 citations
- UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-ClassesTed de Vries Lentsch, Holger Caesar, Dariu GavrilaNeurIPS 2024 · 30 citations
- Pseudo Label Refinery for Unsupervised Domain Adaptation on Cross-Dataset 3D Object DetectionZhanwei Zhang, Minghao Chen, Shuai Xiao, Liang Peng et al.CVPR 2024 · 10 citations
