Sparse Fuse Dense: Towards High Quality 3D Detection with Depth Completion
Xiaopei Wu, Liang Peng, Honghui Yang, Liang Xie, Chenxi Huang, Chengqi Deng, Haifeng Liu, Deng Cai
Abstract
Current LiDAR-only 3D detection methods inevitably suffer from the sparsity of point clouds. Many multi-modal methods are proposed to alleviate this issue, while different representations of images and point clouds make it difficult to fuse them, resulting in suboptimal performance. In this paper, we present a novel multi-modal framework SFD (Sparse Fuse Dense), which utilizes pseudo point clouds generated from depth completion to tackle the issues mentioned above. Different from prior works, we propose a new RoI fusion strategy 3D-GAF (3D Grid-wise Attentive Fusion) to make fuller use of information from different types of point clouds. Specifically, 3D-GAF fuses 3D RoI features from the pair of point clouds in a grid-wise attentive way, which is more fine- grained and more precise. In addition, we propose a SynAugment (Synchronized Augmentation) to enable our multi-modal framework to utilize all data augmentation approaches tailored to LiDAR-only methods. Lastly, we customize an effective and efficient feature extractor CPConv (Color Point Convolution) for pseudo point clouds. It can explore 2D image features and 3D geometric features of pseudo point clouds simultaneously. Our method holds the highest entry on the KITTI car 3D object detection leaderboard <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">†</sup> †On the date of CVPR deadline, i.e., Nov.16, 2021, demonstrating the effectiveness of our SFD. Code will be made publicly available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c5c2d504-2be5-4d24-ba5d-016be7a4c707Cited by top-tier papers31
- Transformation-Equivariant 3D Object Detection for Autonomous DrivingHai Wu, Chenglu Wen, Wei Li, Xin Li et al.AAAI 2023 · 158 citations
- GraphAlign: Enhancing Accurate Feature Alignment by Graph matching for Multi-Modal 3D Object DetectionZiying Song, Haiyue Wei, Lin Bai, Lei Yang et al.ICCV 2023 · 73 citations
- ObjectFusion: Multi-modal 3D Object Detection with Object-Centric FusionQi Cai, Yingwei Pan, Ting Yao, Chong-Wah Ngo et al.ICCV 2023 · 71 citations
- MonoUNI: A Unified Vehicle and Infrastructure-side Monocular 3D Object Detection Network with Sufficient Depth CluesJinrang Jia, Zhenjia Li, Yifeng ShiNeurIPS 2023 · 69 citations
- TransIFF: An Instance-Level Feature Fusion Framework for Vehicle-Infrastructure Cooperative 3D Detection with TransformersZiming Chen, Yifeng Shi, Jinrang JiaICCV 2023 · 55 citations
Builds on23
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou et al.AAAI 2021 · 1,128 citations
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen et al.ICCV 2019 · 840 citations
- Voxel Transformer for 3D Object DetectionJiageng Mao, Yujing Xue, Minzhe Niu, Haoyue Bai et al.ICCV 2021 · 535 citations
- Fast Point R-CNNYilun Chen, Shu Liu, Xiaoyong Shen, Jiaya JiaICCV 2019 · 440 citations
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg et al.ICLR 2020 · 439 citations
Related papers
- Sparse Query Dense: Enhancing 3D Object Detection with Pseudo PointsYujian Mo, Yan Wu, Junqiao Zhao, Zhenjie Hou et al.ACM MM 2024 · 17 citations
- ViKIENet: Towards Efficient 3D Object Detection with Virtual Key Instance Enhanced NetworkZhuochen Yu, Bijie Qiu, Andy W. H. KhongCVPR 2025
- PI-RCNN: An Efficient Multi-Sensor 3D Object Detector with Point-Based Attentive Cont-Conv Fusion ModuleLiang Xie, Chao Xiang, Zhengxu Yu, Guodong Xu et al.AAAI 2020 · 240 citations
- Pseudo-Stereo for Monocular 3D Object Detection in Autonomous DrivingYi-Nan Chen, Hang Dai, Yong DingCVPR 2022 · 91 citations
- Virtual Sparse Convolution for Multimodal 3D Object DetectionHai Wu, Chenglu Wen, Shaoshuai Shi, Xin Li et al.CVPR 2023
