Seg2Box: 3D Object Detection by Point-Wise Semantics Supervision
Maoji Zheng, Ziyu Xu, Qiming Xia, Hai Wu, Chenglu Wen, Cheng Wang
Abstract
LIDAR-based 3D object detection and semantic segmentation are critical tasks in 3D scene understanding. Traditional detection and segmentation methods supervise their models through bounding box labels and semantic mask labels. However, these two independent labels inherently contain significant redundancy. This paper aims to eliminate the redundancy by supervising 3D object detection using only semantic labels. However, the challenge arises due to the incomplete geometry structure and boundary ambiguity of point cloud instances, leading to inaccurate pseudo-labels and poor detection results. To address these challenges, we propose a novel method, named Seg2Box. We first introduce a Multi-Frame Multi-Scale Clustering (MFMS-C) module, which leverages the spatio-temporal consistency of point clouds to generate accurate box-level pseudo-labels. Additionally, the Semantic-Guiding Iterative-Mining Self-Training (SGIM-ST) module is proposed to enhance the performance by progressively refining the pseudo-labels and mining the instances without generating pseudo-labels. Experiments on the Waymo Open Dataset and nuScenes Dataset show that our method significantly outperforms other competitive methods by 23.7% and 10.3% in mAP, respectively. The results demonstrate the great label-efficient potential and advancement of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 21e3305b-d801-4d74-977b-aceb6c6572caCited by top-tier papers2
- Hybrid Robust Collaborative Perception with LiDAR-4D Radar Fusion under Adverse Weather ConditionsYuquan Yang, Hui Zhang, Wenyu Lu, Ziyin Zhang et al.CVPR 2026 · 2 citations
- OWL: Unsupervised 3D Object Detection by Occupancy Guided Warm-up and Large Model Priors ReasoningXusheng Guo, Wanfa Zhang, Shijia Zhao, Qiming Xia et al.AAAI 2026
Builds on12
- Weakly Supervised 3D Object Detection from Point CloudsZengyi Qin, Jinglu Wang, Yan LuACM MM 2020 · 68 citations
- CoIn: Contrastive Instance Feature Mining for Outdoor 3D Object Detection with Very Limited AnnotationsQiming Xia, Jinhao Deng, Chenglu Wen, Hai Wu et al.ICCV 2023 · 34 citations
- Learning to Detect Mobile Objects from LiDAR Scans Without LabelsYurong You, Katie Luo, Cheng Perng Phoo, Wei-Lun Chao et al.CVPR 2022 · 33 citations
- SPGroup3D: Superpoint Grouping Network for Indoor 3D Object DetectionYun Zhu, Le Hui, Yaqi Shen, Jin XieAAAI 2024 · 24 citations
- HINTED: Hard Instance Enhanced Detector with Mixed-Density Feature Fusion for Sparsely-Supervised 3D Object DetectionQiming Xia, Wei Ye, Hai Wu, Shijia Zhao et al.CVPR 2024 · 22 citations
Related papers
- MixSup: Mixed-grained Supervision for Label-efficient LiDAR-based 3D Object DetectionYuxue Yang, Lue Fan, Zhaoxiang ZhangICLR 2024 · 11 citations
- Exploring Geometry-aware Contrast and Clustering Harmonization for Self-supervised 3D Object DetectionHanxue Liang, Chenhan Jiang, Dapeng Feng, Xin Chen et al.ICCV 2021 · 85 citations
- LWSIS: LiDAR-Guided Weakly Supervised Instance Segmentation for Autonomous DrivingXiang Li, Junbo Yin, Botian Shi, Yikang Li et al.AAAI 2023 · 16 citations
- MWSIS: Multimodal Weakly Supervised Instance Segmentation with 2D Box Annotations for Autonomous DrivingGuangfeng Jiang, Jun Liu, Yuzhi Wu, Wenlong Liao et al.AAAI 2024 · 11 citations
- SeMoLi: What Moves Together Belongs TogetherJenny Seidenschwarz, Aljosa Osep, Francesco Ferroni, Simon Lucey et al.CVPR 2024 · 3 citations
