Learning Class Prototypes for Unified Sparse-Supervised 3D Object Detection
Yun Zhu, Le Hui, Hang Yang, Jianjun Qian, Jin Xie, Jian Yang
摘要
Both indoor and outdoor scene perceptions are essential for embodied intelligence. However, current sparse supervised 3D object detection methods focus solely on outdoor scenes without considering indoor settings. To this end, we propose a unified sparse supervised 3D object detection method for both indoor and outdoor scenes through learning class prototypes to effectively utilize unlabeled objects. Specifically, we first propose a prototype-based object mining module that converts the unlabeled object mining into a matching problem between class prototypes and unlabeled features. By using optimal transport matching results, we assign prototype labels to high-confidence features, thereby achieving the mining of unlabeled objects. We then present a multi-label cooperative refinement module to effectively recover missed detections through pseudo label quality control and prototype label cooperation. Experiments show that our method achieves state-of-the-art performance under the one object per scene sparse supervised setting across indoor and outdoor datasets. With only one labeled object per scene, our method achieves about 78%, 90%, and 96% performance compared to the fully supervised detector on ScanNet V2, SUN RGB-D, and KITTI, respectively, highlighting the scalability of our method. Code is available at https://github.com/zyrant/CPDet3D .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- FUSER: Feed-Forward Multiview 3D Registration Transformer and SE(3)^N Diffusion RefinementHaobo Jiang, Jin Xie, Jian Yang, Liang Yu 等CVPR 2026 · 被引用 5 次
- CCF: Complementary Collaborative Fusion for Domain Generalized Multi-Modal 3D Object DetectionYuchen Wu, Kun Wang, Yining Pan, Na ZhaoCVPR 2026 · 被引用 4 次
- Few-Shot Incremental 3D Object Detection in Dynamic Indoor EnvironmentsYun Zhu, Jianjun Qian, Jian Yang, Jin Xie 等CVPR 2026 · 被引用 2 次
- MonoSAOD: Monocular 3D Object Detection with Sparsely Annotated LabelJunyoung Jung, Seokwon Kim, Jung Uk KimCVPR 2026 · 被引用 1 次
- GEM: Generating LiDAR World Model via Deformable MambaYang Wu, Zhaojiang Liu, Qiang Meng, Youquan Liu 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper29
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou 等AAAI 2021 · 被引用 1,128 次
- Exploring Cross-Image Pixel Contrast for Semantic SegmentationWenguan Wang, Tianfei Zhou, Fisher Yu, Jifeng Dai 等ICCV 2021 · 被引用 568 次
- Rethinking Semantic Segmentation: A Prototype ViewTianfei Zhou, Wenguan Wang, Ender Konukoglu, Luc Van GoolCVPR 2022 · 被引用 353 次
- CAGroup3D: Class-Aware Grouping for 3D Object Detection on Point CloudsHaiyang Wang, Lihe Ding, Shaocong Dong, Shaoshuai Shi 等NeurIPS 2022 · 被引用 110 次
相关 Paper
- Commonsense Prototype for Outdoor Unsupervised 3D Object DetectionHai Wu, Shijia Zhao, Xun Huang, Chenglu Wen 等CVPR 2024 · 被引用 15 次
- SS3D: Sparsely-Supervised 3D Object Detection from Point CloudChuandong Liu, Chenqiang Gao, Fangcen Liu, Jiang Liu 等CVPR 2022 · 被引用 32 次
- 3DIoUMatch: Leveraging IoU Prediction for Semi-Supervised 3D Object DetectionHe Wang, Yezhen Cong, Or Litany, Yue Gao 等CVPR 2021
- UniDet3D: Multi-dataset Indoor 3D Object DetectionMaksim Kolodiazhnyi, Anna Vorontsova, Matvey Skripkin, Danila Rukhovich 等AAAI 2025 · 被引用 7 次
- Self-Supervised Pretraining of 3D Features on any Point-CloudZaiwei Zhang, Rohit Girdhar, Armand Joulin, Ishan MisraICCV 2021 · 被引用 333 次
