Eliminating Spatial Ambiguity for Weakly Supervised 3D Object Detection without Spatial Labels
Haizhuang Liu, Huimin Ma, Yilin Wang, Bochao Zou, Tianyu Hu, Rongquan Wang, Jiansheng Chen
摘要
Previous weakly-supervised methods of 3D object detection in driving scenes mainly rely on spatial labels, which provide the location, dimension, or orientation information. The annotation of 3D spatial labels is time-consuming. There also exist methods that do not require spatial labels, but their detections may fall on object parts rather than entire objects or backgrounds. In this paper, a novel cross-modal weakly-supervised 3D progressive refinement framework (WS3DPR) for 3D object detection that only needs image-level class annotations is introduced. The proposed framework consists of two stages: 1) classification refinement for potential objects localization and 2) regression refinement for spatial pseudo labels reasoning. In the first stage, a region proposal network is trained by cross-modal class knowledge transferred from 2D image to 3D point cloud and class information propagation. In the second stage, the locations, dimensions, and orientations of 3D bounding boxes are further refined with geometric reasoning based on 2D frustum and 3D region. When only image-level class labels are available, proposals with different 3D locations become overlapped in 2D, leading to the misclassification of foreground objects. Therefore, a 2D-3D semantic consistency block is proposed to disentangle different 3D proposals after projection. The overall framework progressively learns features in a coarse to fine manner. Comprehensive experiments on the KITTI3D dataset demonstrate that our method achieves competitive performance compared with previous methods with a lightweight labeling process.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Weakly Supervised 3D Object Detection from Point CloudsZengyi Qin, Jinglu Wang, Yan LuACM MM 2020 · 被引用 68 次
- Transferable Semi-Supervised 3D Object Detection From RGB-D DataYew Siang Tang, Gim Hee LeeICCV 2019 · 被引用 41 次
- MWSIS: Multimodal Weakly Supervised Instance Segmentation with 2D Box Annotations for Autonomous DrivingGuangfeng Jiang, Jun Liu, Yuzhi Wu, Wenlong Liao 等AAAI 2024 · 被引用 11 次
- Leveraging Imagery Data with Spatial Point Prior for Weakly Semi-supervised 3D Object DetectionHongzhi Gao, Zheng Chen, Zehui Chen, Lin Chen 等AAAI 2024 · 被引用 3 次
- Weakly Supervised Monocular 3D Object Detection Using Multi-View Projection and Direction ConsistencyRunzhou Tao, Wencheng Han, Zhongying Qiu, Cheng-Zhong Xu 等CVPR 2023
