IDA-3D: Instance-Depth-Aware 3D Object Detection From Stereo Vision for Autonomous Driving
Wanli Peng, Hao Pan, He Liu, Yi Sun
摘要
3D object detection is an important scene understanding task in autonomous driving and virtual reality. Approaches based on LiDAR technology have high performance, but Li-DAR is expensive. Considering more general scenes, where there is no LiDAR data in the 3D datasets, we propose a 3D object detection approach from stereo vision which does not rely on LiDAR data either as input or as supervision in training, but solely takes RGB images with corresponding annotated 3D bounding boxes as training data. As depth estimation of object is the key factor affecting the performance of 3D object detection, we introduce an Instance-Depth-Aware (IDA) module which accurately predicts the depth of the 3D bounding box's center by instance-depth awareness, disparity adaptation and matching cost reweighting. Moreover, our model is an end-to-end learning framework which does not require multiple stages or postprocessing algorithm. We provide detailed experiments on KITTI benchmark and achieve impressive improvements compared with the existing image-based methods. Our code is available at https://github.com/swords123/IDA-3D .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D DetectorXiaoyang Guo, Shaoshuai Shi, Xiaogang Wang, Hongsheng LiICCV 2021 · 被引用 132 次
- Discrete Cosine Transform Network for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Shuang Xu, Zudi Lin 等CVPR 2022 · 被引用 120 次
- Spherical Space Feature Decomposition for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Xiang Gu, Chengli Tan 等ICCV 2023 · 被引用 55 次
- Metric from Human: Zero-shot Monocular Metric Depth Estimation via Test-time AdaptationYizhou Zhao, Hengwei Bian, Kaihua Chen, Pengliang Ji 等NeurIPS 2024 · 被引用 16 次
- Stereo Neural Vernier CaliperShichao Li, Zechun Liu, Zhiqiang Shen, Kwang-Ting ChengAAAI 2022 · 被引用 6 次
它引用的顶会 Paper3
- M3D-RPN: Monocular 3D Region Proposal Network for Object DetectionGarrick Brazil, Xiaoming LiuICCV 2019 · 被引用 542 次
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg 等ICLR 2020 · 被引用 439 次
- Accurate Monocular 3D Object Detection via Color-Embedded 3D Reconstruction for Autonomous DrivingXinzhu Ma, Zhihui Wang, Haojie Li, Pengbo Zhang 等ICCV 2019 · 被引用 339 次
相关 Paper
- Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity EstimationJiaming Sun, Linghao Chen, Yiming Xie, Siyu Zhang 等CVPR 2020
- DSGN: Deep Stereo Geometry Network for 3D Object DetectionYilun Chen, Shu Liu, Xiaoyong Shen, Jiaya JiaCVPR 2020
- VSRD: Instance-Aware Volumetric Silhouette Rendering for Weakly Supervised 3D Object DetectionZihua Liu, Hiroki Sakuma, Masatoshi OkutomiCVPR 2024
- Is Pseudo-Lidar needed for Monocular 3D Object detection?Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li 等ICCV 2021 · 被引用 404 次
- Training an Open-Vocabulary Monocular 3D Detection Model without 3D DataRui Huang, Henry Zheng, Yan Wang, Zhuofan Xia 等NeurIPS 2024 · 被引用 26 次
