3DIoUMatch: Leveraging IoU Prediction for Semi-Supervised 3D Object Detection
He Wang, Yezhen Cong, Or Litany, Yue Gao, Leonidas J. Guibas
Abstract
3D object detection is an important yet demanding task that heavily relies on difficult to obtain 3D annotations. To reduce the required amount of supervision, we propose 3DIoUMatch, a novel semi-supervised method for 3D object detection applicable to both indoor and outdoor scenes. We leverage a teacher-student mutual learning framework to propagate information from the labeled to the unlabeled train set in the form of pseudo-labels. However, due to the high task complexity, we observe that the pseudo-labels suffer from significant noise and are thus not directly usable. To that end, we introduce a confidence-based filtering mechanism, inspired by FixMatch. We set confidence thresholds based upon the predicted objectness and class probability to filter low-quality pseudo-labels. While effective, we observe that these two measures do not sufficiently capture localization quality. We therefore propose to use the estimated 3D IoU as a localization metric and set category-aware selfadjusted thresholds to filter poorly localized proposals. We adopt VoteNet as our backbone detector on indoor datasets while we use PV-RCNN on the autonomous driving dataset, KITTI. Our method consistently improves state-of-the-art methods on both ScanNet and SUN-RGBD benchmarks by significant margins under all label ratios (including fully labeled setting). For example, when training using only 10% labeled data on ScanNet, 3DIoUMatch achieves 7.7 absolute improvement on mAP@0.25 and 8.5 absolute improvement on mAP@0.5 upon the prior art. On KITTI, we are the first to demonstrate semi-supervised 3D object detection and our method surpasses a fully supervised baseline from 1.8% to 7.6% under different label ratio and categories.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9844207c-3f80-4c72-8df3-a79425b4b34bCited by top-tier papers34
- Guided Point Contrastive Learning for Semi-supervised Point Cloud Semantic SegmentationLi Jiang, Shaoshuai Shi, Zhuotao Tian, Xin Lai et al.ICCV 2021 · 137 citations
- Learning from Temporal Gradient for Semi-supervised Action RecognitionJunfei Xiao, Longlong Jing, Lin Zhang, Ju He et al.CVPR 2022 · 77 citations
- SSDA3D: Semi-supervised Domain Adaptation for 3D Object Detection from Point CloudYan Wang, Junbo Yin, Wei Li, Pascal Frossard et al.AAAI 2023 · 60 citations
- A Simple Vision Transformer for Weakly Semi-supervised 3D Object DetectionDingyuan Zhang, Dingkang Liang, Zhikang Zou, Jingyu Li et al.ICCV 2023 · 36 citations
- Combating Noise: Semi-supervised Learning by Region Uncertainty QuantificationZhenyu Wang, Ya-Li Li, Ye Guo, Shengjin WangNeurIPS 2021 · 34 citations
Builds on10
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
Related papers
- Transferable Semi-Supervised 3D Object Detection From RGB-D DataYew Siang Tang, Gim Hee LeeICCV 2019 · 41 citations
- Diffusion-SS3D: Diffusion Model for Semi-supervised 3D Object DetectionCheng-Ju Ho, Chen-Hsuan Tai, Yen-Yu Lin, Ming-Hsuan Yang et al.NeurIPS 2023 · 31 citations
- Dual-Perspective Knowledge Enrichment for Semi-supervised 3D Object DetectionYucheng Han, Na Zhao, Weiling Chen, Keng Teck Ma et al.AAAI 2024 · 11 citations
- DQS3D: Densely-matched Quantization-aware Semi-supervised 3D DetectionHuan-ang Gao, Beiwen Tian, Pengfei Li, Hao Zhao et al.ICCV 2023 · 21 citations
- Learning Class Prototypes for Unified Sparse-Supervised 3D Object DetectionYun Zhu, Le Hui, Hang Yang, Jianjun Qian et al.CVPR 2025
