Learning Depth-Guided Convolutions for Monocular 3D Object Detection
Mingyu Ding, Yuqi Huo, Hongwei Yi, Zhe Wang, Jianping Shi, Zhiwu Lu, Ping Luo
Abstract
3D object detection from a single image without LiDAR is a challenging task due to the lack of accurate depth information. Conventional 2D convolutions are unsuitable for this task because they fail to capture local object and its scale information, which are vital for 3D object detection. To better represent 3D structure, prior arts typically transform depth maps estimated from 2D images into a pseudo-LiDAR representation, and then apply existing 3D point-cloud based object detectors. However, their results depend heavily on the accuracy of the estimated depth maps, resulting in suboptimal performance. In this work, instead of using pseudo-LiDAR representation, we improve the fundamental 2D fully convolutions by proposing a new local convolutional network (LCN), termed Depth-guided Dynamic-Depthwise-Dilated LCN (D 4 LCN), where the filters and their receptive fields can be automatically learned from image-based depth maps, making different pixels of different images have different filters. D 4 LCN overcomes the limitation of conventional 2D convolutions and narrows the gap between image representation and 3D point cloud representation. Extensive experiments show that D 4 LCN outperforms existing works by large margins. For example, the relative improvement of D 4 LCN against the state-of-theart on KITTI is 9.1% in the moderate setting. D 4 LCN ranks 1 st on KITTI monocular 3D object detection benchmark at the time of submission (car, December 2019). The code is available at https://github.com/dingmyu/D4LCN .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 720af8c6-a2a2-4ccb-befa-cb3dc0e7242dCited by top-tier papers73
- Is Pseudo-Lidar needed for Monocular 3D Object detection?Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li et al.ICCV 2021 · 404 citations
- Geometry Uncertainty Projection Network for Monocular 3D Object DetectionYan Lu, Xinzhu Ma, Lei Yang, Tianzhu Zhang et al.ICCV 2021 · 294 citations
- PolarFormer: Multi-Camera 3D Object Detection with Polar TransformerYanqin Jiang, Li Zhang, Zhenwei Miao, Xiatian Zhu et al.AAAI 2023 · 240 citations
- BEVStereo: Enhancing Depth Estimation in Multi-View 3D Object Detection with Temporal StereoYinhao Li, Han Bao, Zheng Ge, Jinrong Yang et al.AAAI 2023 · 226 citations
- MonoDTR: Monocular 3D Object Detection with Depth-Aware TransformerKuan-Chih Huang, Tsung-Han Wu, Hung-Ting Su, Winston H. HsuCVPR 2022 · 199 citations
Builds on6
- Scale-Aware Trident Networks for Object DetectionYanghao Li, Yuntao Chen, Naiyan Wang, Zhaoxiang ZhangICCV 2019 · 1,031 citations
- M3D-RPN: Monocular 3D Region Proposal Network for Object DetectionGarrick Brazil, Xiaoming LiuICCV 2019 · 542 citations
- Disentangling Monocular 3D Object DetectionAndrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Manuel Lopez-Antequera et al.ICCV 2019 · 504 citations
- Fast Point R-CNNYilun Chen, Shu Liu, Xiaoyong Shen, Jiaya JiaICCV 2019 · 440 citations
- Accurate Monocular 3D Object Detection via Color-Embedded 3D Reconstruction for Autonomous DrivingXinzhu Ma, Zhihui Wang, Haojie Li, Pengbo Zhang et al.ICCV 2019 · 339 citations
Related papers
- Pseudo-Stereo for Monocular 3D Object Detection in Autonomous DrivingYi-Nan Chen, Hang Dai, Yong DingCVPR 2022 · 91 citations
- Depth-Conditioned Dynamic Message Propagation for Monocular 3D Object DetectionLi Wang, Liang Du, Xiaoqing Ye, Yanwei Fu et al.CVPR 2021
- Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity EstimationJiaming Sun, Linghao Chen, Yiming Xie, Siyu Zhang et al.CVPR 2020
- End-to-End Pseudo-LiDAR for Image-Based 3D Object DetectionRui Qian, Divyansh Garg, Yan Wang, Yurong You et al.CVPR 2020
- IDA-3D: Instance-Depth-Aware 3D Object Detection From Stereo Vision for Autonomous DrivingWanli Peng, Hao Pan, He Liu, Yi SunCVPR 2020
