DOPS: Learning to Detect 3D Objects and Predict Their 3D Shapes
Mahyar Najibi, Guangda Lai, Abhijit Kundu, Zhichao Lu, Vivek Rathod, Thomas A. Funkhouser, Caroline Pantofaru, David A. Ross, Larry S. Davis, Alireza Fathi
Abstract
We propose DOPS, a fast single-stage 3D object detection method for LIDAR data. Previous methods often make domain-specific design decisions, for example projecting points into a bird-eye view image in autonomous driving scenarios. In contrast, we propose a general-purpose method that works on both indoor and outdoor scenes. The core novelty of our method is a fast, single-pass architecture that both detects objects in 3D and estimates their shapes. 3D bounding box parameters are estimated in one pass for every point, aggregated through graph convolutions, and fed into a branch of the network that predicts latent codes representing the shape of each detected object. The latent shape space and shape decoder are learned on a synthetic dataset and then used as supervision for the end-toend training of the 3D object detection pipeline. Thus our model is able to extract shapes without access to groundtruth shape information in the target dataset. During experiments, we find that our proposed method achieves stateof-the-art results by ∼5% on object detection in ScanNet scenes, and it gets top results by 3.4% in the Waymo Open Dataset, while reproducing the shapes of detected cars.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bbcd03b6-efa3-4374-ab01-1b9968681996Cited by top-tier papers15
- Behind the Curtain: Learning Occluded Shapes for 3D Object DetectionQiangeng Xu, Yiqi Zhong, Ulrich NeumannAAAI 2022 · 188 citations
- Object DGCNN: 3D Object Detection using Dynamic GraphsYue Wang, Justin M. SolomonNeurIPS 2021 · 127 citations
- VENet: Voting Enhancement Network for 3D Object DetectionQian Xie, Yu-Kun Lai, Jing Wu, Zhoutao Wang et al.ICCV 2021 · 60 citations
- Unsupervised 3D Perception with 2D Vision-Language Distillation for Autonomous DrivingMahyar Najibi, Jingwei Ji, Yin Zhou, Charles R. Qi et al.ICCV 2023 · 52 citations
- FWD: Real-time Novel View Synthesis with Forward Warping and DepthAng Cao, Chris Rockwell, Justin JohnsonCVPR 2022 · 45 citations
Builds on3
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- ShapeMask: Learning to Segment Novel Objects by Refining Shape PriorsWeicheng Kuo, Anelia Angelova, Jitendra Malik, Tsung-Yi LinICCV 2019 · 127 citations
Related papers
- Sparse2Dense: Learning to Densify 3D Features for 3D Object DetectionTianyu Wang, Xiaowei Hu, Zhengzhe Liu, Chi-Wing FuNeurIPS 2022 · 24 citations
- Learning 3D Scene Priors with 2D SupervisionYinyu Nie, Angela Dai, Xiaoguang Han, Matthias NießnerCVPR 2023
- Fully Convolutional One-Stage 3D Object Detection on LiDAR Range ImagesZhi Tian, Xiangxiang Chu, Xiaoming Wang, Xiaolin Wei et al.NeurIPS 2022 · 168 citations
- RangeDet: In Defense of Range View for LiDAR-based 3D Object DetectionLue Fan, Xuan Xiong, Feng Wang, Naiyan Wang et al.ICCV 2021 · 268 citations
- SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point CloudsQingdong He, Zhengning Wang, Hao Zeng, Yi Zeng et al.AAAI 2022 · 124 citations
