Semi-Supervised Stereo-Based 3D Object Detection via Cross-View Consensus
Wenhao Wu, Hau-San Wong, Si Wu
摘要
Stereo-based 3D object detection, which aims at detecting 3D objects with stereo cameras, shows great potential in low-cost deployment compared to LiDAR-based methods and excellent performance compared to monocular-based algorithms. However, the impressive performance of stereobased 3D object detection is at the huge cost of high-quality manual annotations, which are hardly attainable for any given scene. Semi-supervised learning, in which limited annotated data and numerous unannotated data are required to achieve a satisfactory model, is a promising method to address the problem of data deficiency. In this work, we propose to achieve semi-supervised learning for stereo-based 3D object detection through pseudo annotation generation from a temporal-aggregated teacher model, which temporally accumulates knowledge from a student model. To facilitate a more stable and accurate depth estimation, we introduce Temporal-Aggregation-Guided (TAG) disparity consistency, a cross-view disparity consistency constraint between the teacher model and the student model for robust and improved depth estimation. To mitigate noise in pseudo annotation generation, we propose a cross-view agreement strategy, in which pseudo annotations should attain high degree of agreements between 3D and 2D views, as well as between binocular views. We perform extensive experiments on the KITTI 3D dataset to demonstrate our proposed method's capability in leveraging a huge amount of unannotated stereo images to attain significantly improved detection results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Real-time Stereo-based 3D Object Detection for Streaming PerceptionChangcai Li, Zonghua Gu, Gang Chen, Libo Huang 等NeurIPS 2024 · 被引用 3 次
- SPAN: Spatial-Projection Alignment for Monocular 3D Object DetectionYifan Wang, Yian Zhao, Fanqi Pu, Xiaochen Yang 等CVPR 2026
- Single-to-Dual-View Adaptation for Egocentric 3D Hand Pose EstimationRuicong Liu, Takehiko Ohkawa, Mingfang Zhang, Yoichi SatoCVPR 2024
它引用的顶会 Paper24
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- End-to-End Semi-Supervised Object Detection with Soft TeacherMengde Xu, Zheng Zhang, Han Hu, Jianfeng Wang 等ICCV 2021 · 被引用 622 次
- Unbiased Teacher for Semi-Supervised Object DetectionYen-Cheng Liu, Chih-Yao Ma, Zijian He, Chia-Wen Kuo 等ICLR 2021 · 被引用 603 次
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg 等ICLR 2020 · 被引用 439 次
- LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D DetectorXiaoyang Guo, Shaoshuai Shi, Xiaogang Wang, Hongsheng LiICCV 2021 · 被引用 132 次
相关 Paper
- Learning with Noisy Data for Semi-Supervised 3D Object DetectionZehui Chen, Zhenyu Li, Shuo Wang, Dengpan Fu 等ICCV 2023 · 被引用 14 次
- SS3D: Sparsely-Supervised 3D Object Detection from Point CloudChuandong Liu, Chenqiang Gao, Fangcen Liu, Jiang Liu 等CVPR 2022 · 被引用 32 次
- Decoupled Pseudo-Labeling for Semi-Supervised Monocular 3D Object DetectionJiacheng Zhang, Jiaming Li, Xiangru Lin, Wei Zhang 等CVPR 2024 · 被引用 17 次
- Leveraging Imagery Data with Spatial Point Prior for Weakly Semi-supervised 3D Object DetectionHongzhi Gao, Zheng Chen, Zehui Chen, Lin Chen 等AAAI 2024 · 被引用 3 次
- 3DIoUMatch: Leveraging IoU Prediction for Semi-Supervised 3D Object DetectionHe Wang, Yezhen Cong, Or Litany, Yue Gao 等CVPR 2021
