Pyramid R-CNN: Towards Better Performance and Adaptability for 3D Object Detection
Jiageng Mao, Minzhe Niu, Haoyue Bai, Xiaodan Liang, Hang Xu, Chunjing Xu
摘要
We present a flexible and high-performance framework, named Pyramid R-CNN, for two-stage 3D object detection from point clouds. Current approaches generally rely on the points or voxels of interest for RoI feature extraction on the second stage, but cannot effectively handle the sparsity and non-uniform distribution of those points, and this may result in failures in detecting objects that are far away. To resolve the problems, we propose a novel second-stage module, named pyramid RoI head, to adaptively learn the features from the sparse points of interest. The pyramid RoI head consists of three key components. Firstly, we propose the RoI-grid Pyramid, which mitigates the sparsity problem by extensively collecting points of interest for each RoI in a pyramid manner. Secondly, we propose RoI-grid Attention, a new operation that can encode richer information from sparse points by incorporating conventional attention-based and graph-based point operators into a unified formulation. Thirdly, we propose the Density-Aware Radius Prediction (DARP) module, which can adapt to different point density levels by dynamically adjusting the focusing range of RoIs. Combining the three components, our pyramid RoI head is robust to the sparse and imbalanced circumstances, and can be applied upon various 3D backbones to consistently boost the detection performance. Extensive experiments show that Pyramid R-CNN outperforms the state-of-the-art 3D detection models by a large margin on both the KITTI dataset and the Waymo Open dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point CloudsYifan Zhang, Qingyong Hu, Guoquan Xu, Yanxin Ma 等CVPR 2022 · 被引用 376 次
- Focal Sparse Convolutional Networks for 3D Object DetectionYukang Chen, Yanwei Li, Xiangyu Zhang, Jian Sun 等CVPR 2022 · 被引用 293 次
- Sparse Fuse Dense: Towards High Quality 3D Detection with Depth CompletionXiaopei Wu, Liang Peng, Honghui Yang, Liang Xie 等CVPR 2022 · 被引用 248 次
- Fully Sparse 3D Object DetectionLue Fan, Feng Wang, Naiyan Wang, Zhaoxiang ZhangNeurIPS 2022 · 被引用 168 次
- RNNPose: Recurrent 6-DoF Object Pose Refinement with Robust Correspondence Field Estimation and Pose OptimizationYan Xu, Kwan-Yee Lin, Guofeng Zhang, Xiaogang Wang 等CVPR 2022 · 被引用 82 次
它引用的顶会 Paper12
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou 等AAAI 2021 · 被引用 1,128 次
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen 等ICCV 2019 · 被引用 840 次
- TANet: Robust 3D Object Detection from Point Clouds with Triple AttentionZhe Liu, Xin Zhao, Tengteng Huang, Ruolan Hu 等AAAI 2020 · 被引用 412 次
- Interpolated Convolutional Networks for 3D Point Cloud UnderstandingJiageng Mao, Xiaogang Wang, Hongsheng LiICCV 2019 · 被引用 241 次
- Every View Counts: Cross-View Consistency in 3D Object Detection with Hybrid-Cylindrical-Spherical VoxelizationQi Chen, Lin Sun, Ernest Cheung, Alan L. YuilleNeurIPS 2020 · 被引用 124 次
相关 Paper
- PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object DetectionShaoshuai Shi, Chaoxu Guo, Li Jiang, Zhe Wang 等CVPR 2020
- Point Density-Aware Voxels for LiDAR 3D Object DetectionJordan S. K. Hu, Tianshu Kuai, Steven L. WaslanderCVPR 2022
- Fast Point R-CNNYilun Chen, Shu Liu, Xiaoyong Shen, Jiaya JiaICCV 2019 · 被引用 440 次
- HVPR: Hybrid Voxel-Point Representation for Single-Stage 3D Object DetectionJongyoun Noh, Sanghoon Lee, Bumsub HamCVPR 2021
- CAGroup3D: Class-Aware Grouping for 3D Object Detection on Point CloudsHaiyang Wang, Lihe Ding, Shaocong Dong, Shaoshuai Shi 等NeurIPS 2022 · 被引用 110 次
