Autolabeling 3D Objects With Differentiable Rendering of SDF Shape Priors
Sergey Zakharov, Wadim Kehl, Arjun Bhargava, Adrien Gaidon
Abstract
We present an automatic annotation pipeline to recover 9D cuboids and 3D shapes from pre-trained off-the-shelf 2D detectors and sparse LIDAR data. Our autolabeling method solves an ill-posed inverse problem by considering learned shape priors and optimizing geometric and physical parameters. To address this challenging problem, we apply a novel differentiable shape renderer to signed distance fields (SDF), leveraged together with normalized object coordinate spaces (NOCS). Initially trained on synthetic data to predict shape and coordinates, our method uses these predictions for projective and geometric alignment over real samples. Moreover, we also propose a curriculum learning strategy, iteratively retraining on samples of increasing difficulty in subsequent self-improving annotation rounds. Our experiments on the KITTI3D dataset show that we can recover a substantial amount of accurate cuboids, and that these autolabels can be used to train 3D vehicle detectors with state-of-the-art results. The code is available at github.com/TRI-ML/sdflabel.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 59f08b88-710b-45d6-bc76-709264450571Cited by top-tier papers31
- Neural-Pull: Learning Signed Distance Function from Point clouds by Learning to Pull Space onto SurfaceBaorui Ma, Zhizhong Han, Yu-Shen Liu, Matthias ZwickerICML 2021 · 215 citations
- AutoShape: Real-Time Shape-Aware Monocular 3D Object DetectionZongdai Liu, Dingfu Zhou, Feixiang Lu, Jin Fang et al.ICCV 2021 · 176 citations
- ADOP: approximate differentiable one-pixel point renderingDarius Rückert, Linus Franke, Marc StammingerSIGGRAPH 2022 · 127 citations
- CaSPR: Learning Canonical Spatiotemporal Point Cloud RepresentationsDavis Rempe, Tolga Birdal, Yongheng Zhao, Zan Gojcic et al.NeurIPS 2020 · 78 citations
- DRWR: A Differentiable Renderer without Rendering for Unsupervised 3D Structure Learning from Silhouette ImagesZhizhong Han, Chao Chen, Yu-Shen Liu, Matthias ZwickerICML 2020 · 60 citations
Builds on8
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 789 citations
- Pix2Pose: Pixel-Wise Coordinate Regression of Objects for 6D Pose EstimationKiru Park, Timothy Patten, Markus VinczeICCV 2019 · 527 citations
- Disentangling Monocular 3D Object DetectionAndrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Manuel Lopez-Antequera et al.ICCV 2019 · 504 citations
- CDPN: Coordinates-Based Disentangled Pose Network for Real-Time RGB-Based 6-DoF Object Pose EstimationZhigang Li, Gu Wang, Xiangyang JiICCV 2019 · 482 citations
- Three-D Safari: Learning to Estimate Zebra Pose, Shape, and Texture From Images "In the Wild"Silvia Zuffi, Angjoo Kanazawa, Tanya Y. Berger-Wolf, Michael J. BlackICCV 2019 · 183 citations
Related papers
- VSRD: Instance-Aware Volumetric Silhouette Rendering for Weakly Supervised 3D Object DetectionZihua Liu, Hiroki Sakuma, Masatoshi OkutomiCVPR 2024
- MV-DeepSDF: Implicit Modeling with Multi-Sweep Point Clouds for 3D Vehicle Reconstruction in Autonomous DrivingYibo Liu, Kelly Zhu, Guile Wu, Yuan Ren et al.ICCV 2023 · 18 citations
- NeurOCS: Neural NOCS Supervision for Monocular 3D Object LocalizationZhixiang Min, Bingbing Zhuang, Samuel Schulter, Buyu Liu et al.CVPR 2023
- Image-to-Lidar Self-Supervised Distillation for Autonomous Driving DataCorentin Sautier, Gilles Puy, Spyros Gidaris, Alexandre Boulch et al.CVPR 2022 · 102 citations
- UniPAD: A Universal Pre-Training Paradigm for Autonomous DrivingHonghui Yang, Sha Zhang, Di Huang, Xiaoyang Wu et al.CVPR 2024 · 31 citations
