To the Point: Efficient 3D Object Detection in the Range Image With Graph Convolution Kernels
Yuning Chai, Pei Sun, Jiquan Ngiam, Weiyue Wang, Benjamin Caine, Vijay Vasudevan, Xiao Zhang, Dragomir Anguelov
摘要
3D object detection is vital for many robotics applications. For tasks where a 2D perspective range image exists, we propose to learn a 3D representation directly from this range image view. To this end, we designed a 2D convolutional network architecture that carries the 3D spherical coordinates of each pixel throughout the network. Its layers can consume any arbitrary convolution kernel in place of the default inner product kernel and exploit the underlying local geometry around each pixel. We outline four such kernels: a dense kernel according to the bag-of-words paradigm, and three graph kernels inspired by recent graph neural network advances: the Transformer, the PointNet, and the Edge Convolution. We also explore cross-modality fusion with the camera image, facilitated by operating in the perspective range image view. Our method performs competitively on the Waymo Open Dataset and improves the state-of-the-art AP for pedestrian detection from 69.7% to 75.5%. It is also efficient in that our smallest model, which still outperforms the popular PointPillars in quality, requires 180 times fewer FLOPS and model parameters.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- DeepInteraction: 3D Object Detection via Modality InteractionZeyu Yang, Jiaqi Chen, Zhenwei Miao, Wei Li 等NeurIPS 2022 · 被引用 268 次
- Sparse Fuse Dense: Towards High Quality 3D Detection with Depth CompletionXiaopei Wu, Liang Peng, Honghui Yang, Liang Xie 等CVPR 2022 · 被引用 248 次
- Voxel Set Transformer: A Set-to-Set Approach to 3D Object Detection from Point CloudsChenhang He, Ruihuang Li, Shuai Li, Lei ZhangCVPR 2022 · 被引用 217 次
- AFDetV2: Rethinking the Necessity of the Second Stage for Object Detection from Point CloudsYihan Hu, Zhuangzhuang Ding, Runzhou Ge, Wenxin Shao 等AAAI 2022 · 被引用 163 次
- ObjectFusion: Multi-modal 3D Object Detection with Object-Centric FusionQi Cai, Yingwei Pan, Ting Yao, Chong-Wah Ngo 等ICCV 2023 · 被引用 71 次
它引用的顶会 Paper5
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World DataMikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua, Duc Thanh Nguyen 等ICCV 2019 · 被引用 1,003 次
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora 等CVPR 2020
- PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object DetectionShaoshuai Shi, Chaoxu Guo, Li Jiang, Zhe Wang 等CVPR 2020
- Scalability in Perception for Autonomous Driving: Waymo Open DatasetPei Sun, Henrik Kretzschmar, Xerxes Dotiwalla, Aurelien Chouard 等CVPR 2020
相关 Paper
- RSN: Range Sparse Net for Efficient, Accurate LiDAR 3D Object DetectionPei Sun, Weiyue Wang, Yuning Chai, Gamaleldin Elsayed 等CVPR 2021
- PVT-SSD: Single-Stage 3D Object Detector with Point-Voxel TransformerHonghui Yang, Wenxiao Wang, Minghao Chen, Binbin Lin 等CVPR 2023
- RangeDet: In Defense of Range View for LiDAR-based 3D Object DetectionLue Fan, Xuan Xiong, Feng Wang, Naiyan Wang 等ICCV 2021 · 被引用 268 次
- RangeIoUDet: Range Image Based Real-Time 3D Object Detector Optimized by Intersection Over UnionZhidong Liang, Zehan Zhang, Ming Zhang, Xian Zhao 等CVPR 2021
- RangePerception: Taming LiDAR Range View for Efficient and Accurate 3D Object DetectionYeqi Bai, Ben Fei, Youquan Liu, Tao Ma 等NeurIPS 2023 · 被引用 11 次
