End-to-End Pseudo-LiDAR for Image-Based 3D Object Detection
Rui Qian, Divyansh Garg, Yan Wang, Yurong You, Serge J. Belongie, Bharath Hariharan, Mark E. Campbell, Kilian Q. Weinberger, Wei-Lun Chao
Abstract
Reliable and accurate 3D object detection is a necessity for safe autonomous driving. Although LiDAR sensors can provide accurate 3D point cloud estimates of the environment, they are also prohibitively expensive for many settings. Recently, the introduction of pseudo-LiDAR (PL) has led to a drastic reduction in the accuracy gap between methods based on LiDAR sensors and those based on cheap stereo cameras. PL combines state-of-the-art deep neural networks for 3D depth estimation with those for 3D object detection by converting 2D depth map outputs to 3D point cloud inputs. However, so far these two networks have to be trained separately. In this paper, we introduce a new framework based on differentiable Change of Representation (CoR) modules that allow the entire PL pipeline to be trained end-to-end. The resulting framework is compatible with most state-of-the-art networks for both tasks and in combination with PointRCNN improves over PL consistently across all benchmarks -yielding the highest entry on the KITTI image-based 3D object detection leaderboard at the time of submission. Our code will be made available at https://github.com/mileyan/ pseudo-LiDAR_e2e .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 478210dd-0625-450d-a1ac-8d8df384ff44Cited by top-tier papers29
- BEVDepth: Acquisition of Reliable Depth for Multi-View 3D Object DetectionYinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang et al.AAAI 2023 · 954 citations
- DeepFusion: Lidar-Camera Deep Fusion for Multi-Modal 3D Object DetectionYingwei Li, Adams Wei Yu, Tianjian Meng, Benjamin Caine et al.CVPR 2022 · 508 citations
- Is Pseudo-Lidar needed for Monocular 3D Object detection?Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li et al.ICCV 2021 · 404 citations
- Multimodal Virtual Point 3D DetectionTianwei Yin, Xingyi Zhou, Philipp KrähenbühlNeurIPS 2021 · 379 citations
- Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point CloudsYifan Zhang, Qingyong Hu, Guoquan Xu, Yanxin Ma et al.CVPR 2022 · 376 citations
Builds on5
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen et al.ICCV 2019 · 840 citations
- Fast Point R-CNNYilun Chen, Shu Liu, Xiaoyong Shen, Jiaya JiaICCV 2019 · 440 citations
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg et al.ICLR 2020 · 439 citations
- ZoomNet: Part-Aware Adaptive Zooming Neural Network for 3D Object DetectionZhenbo Xu, Wei Zhang, Xiaoqing Ye, Xiao Tan et al.AAAI 2020 · 77 citations
- Train in Germany, Test in the USA: Making 3D Object Detectors GeneralizeYan Wang, Xiangyu Chen, Yurong You, Li Erran Li et al.CVPR 2020
Related papers
- Pseudo-Stereo for Monocular 3D Object Detection in Autonomous DrivingYi-Nan Chen, Hang Dai, Yong DingCVPR 2022 · 91 citations
- DSGN: Deep Stereo Geometry Network for 3D Object DetectionYilun Chen, Shu Liu, Xiaoyong Shen, Jiaya JiaCVPR 2020
- Neighbor-Vote: Improving Monocular 3D Object Detection through Neighbor Distance VotingXiaomeng Chu, Jiajun Deng, Yao Li, Zhenxun Yuan et al.ACM MM 2021 · 24 citations
- Are we Missing Confidence in Pseudo-LiDAR Methods for Monocular 3D Object Detection?Andrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Peter Kontschieder et al.ICCV 2021 · 37 citations
- IDA-3D: Instance-Depth-Aware 3D Object Detection From Stereo Vision for Autonomous DrivingWanli Peng, Hao Pan, He Liu, Yi SunCVPR 2020
