Point Density-Aware Voxels for LiDAR 3D Object Detection
Jordan S. K. Hu, Tianshu Kuai, Steven L. Waslander
Abstract
LiDAR has become one of the primary 3D object detection sensors in autonomous driving. However, LiDAR's diverging point pattern with increasing distance results in a non-uniform sampled point cloud ill-suited to discretized volumetric feature extraction. Current methods either rely on voxelized point clouds or use inefficient farthest point sampling to mitigate detrimental effects caused by density variation but largely ignore point density as a feature and its predictable relationship with distance from the LiDAR sensor. Our proposed solution, Point Density-Aware Voxel network (PDV), is an end-to-end two stage LiDAR 3D object detection architecture that is designed to account for these point density variations. PDV efficiently localizes voxel features from the 3D sparse convolution backbone through voxel point centroids. The spatially localized voxel features are then aggregated through a density-aware RoI grid pooling module using kernel density estimation (KDE) and self attention with point density positional encoding. Finally, we exploit LiDAR's point density to distance relationship to refine our final bounding box confidences. PDV outperforms all state-of-the-art methods on the Waymo Open Dataset and achieves competitive results on the KITTI dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dd501fbd-8e5c-4a14-9576-321bb818e7ddCited by top-tier papers24
- IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object DetectionJunbo Yin, Jianbing Shen, Runnan Chen, Wei Li et al.CVPR 2024 · 73 citations
- PG-RCNN: Semantic Surface Point Generation for 3D Object DetectionInyong Koo, Inyoung Lee, Se-Ho Kim, Hee-Seon Kim et al.ICCV 2023 · 47 citations
- DetZero: Rethinking Offboard 3D Object Detection with Long-term Sequential Point CloudsTao Ma, Xuemeng Yang, Hongbin Zhou, Xin Li et al.ICCV 2023 · 46 citations
- Revisiting Domain-Adaptive 3D Object Detection by Reliable, Diverse and Class-balanced Pseudo-LabelingZhuoxiao Chen, Yadan Luo, Zheng Wang, Mahsa Baktashmotlagh et al.ICCV 2023 · 40 citations
- A Simple Vision Transformer for Weakly Semi-supervised 3D Object DetectionDingyuan Zhang, Dingkang Liang, Zhikang Zou, Jingyu Li et al.ICCV 2023 · 36 citations
Builds on12
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou et al.AAAI 2021 · 1,128 citations
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen et al.ICCV 2019 · 840 citations
- Voxel Transformer for 3D Object DetectionJiageng Mao, Yujing Xue, Minzhe Niu, Haoyue Bai et al.ICCV 2021 · 535 citations
- Improving 3D Object Detection with Channel-wise TransformerHualian Sheng, Sijia Cai, Yuan Liu, Bing Deng et al.ICCV 2021 · 293 citations
Related papers
- SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point CloudsQingdong He, Zhengning Wang, Hao Zeng, Yi Zeng et al.AAAI 2022 · 124 citations
- PVGNet: A Bottom-Up One-Stage 3D Object Detector With Integrated Multi-Level FeaturesZhenwei Miao, Jikai Chen, Hongyu Pan, Ruiwen Zhang et al.CVPR 2021
- Pyramid R-CNN: Towards Better Performance and Adaptability for 3D Object DetectionJiageng Mao, Minzhe Niu, Haoyue Bai, Xiaodan Liang et al.ICCV 2021 · 173 citations
- VoxelKP: A Voxel-Based Network Architecture for Human Keypoint Estimation in LiDAR DataJian Shi, Peter WonkaICCV 2025
- LiDAR R-CNN: An Efficient and Universal 3D Object DetectorZhichao Li, Feng Wang, Naiyan WangCVPR 2021
