RGB-D Local Implicit Function for Depth Completion of Transparent Objects
Luyang Zhu, Arsalan Mousavian, Yu Xiang, Hammad Mazhar, Jozef van Eenbergen, Shoubhik Debnath, Dieter Fox
Abstract
Majority of the perception methods in robotics require depth information provided by RGB-D cameras. However, standard 3D sensors fail to capture depth of transparent objects due to refraction and absorption of light. In this paper, we introduce a new approach for depth completion of transparent objects from a single RGB-D image. Key to our approach is a local implicit neural representation built on ray-voxel pairs that allows our method to generalize to unseen objects and achieve fast inference speed. Based on this representation, we present a novel framework that can complete missing depth given noisy RGB-D input. We further improve the depth estimation iteratively using a self-correcting refinement model. To train the whole pipeline, we build a large scale synthetic dataset with transparent objects. Experiments demonstrate that our method performs significantly better than the current stateof-the-art methods on both synthetic and real world data. In addition, our approach improves the inference speed by a factor of 20 compared to the previous best method, ClearGrasp [43]. Code will be released at https : //research.nvidia.com/publication/2021- 03_RGB-D-Local-Implicit.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50253d5b-5685-45dc-a2d7-df386ded70feCited by top-tier papers11
- Toward Practical Monocular Indoor Depth EstimationCho-Ying Wu, Jialiang Wang, Michael Hall, Ulrich Neumann et al.CVPR 2022 · 68 citations
- Transparent Shape from a Single View Polarization ImageMingqi Shao, Chongkun Xia, Zhendong Yang, Junnan Huang et al.ICCV 2023 · 22 citations
- Consistent Depth Prediction for Transparent Object Reconstruction from RGB-D CameraYuxiang Cai, Yifan Zhu, Haiwei Zhang, Bo RenICCV 2023 · 10 citations
- DepthFocus: Controllable Depth Estimation for See-Through Scenesjunhong min, Jimin Kim, Minwook Kim, Cheol-Hui Min et al.CVPR 2026 · 4 citations
- TransNormal: Dense Visual Semantics for Diffusion-based Transparent Object Normal EstimationMingwei Li, Hehe Fan, Yi YangICML 2026 · 1 citation
Builds on11
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua et al.NeurIPS 2020 · 1,535 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 1,001 citations
- Learning Shape Templates With Structured Implicit FunctionsKyle Genova, Forrester Cole, Daniel Vlasic, Aaron Sarna et al.ICCV 2019 · 427 citations
- Depth Completion From Sparse LiDAR Data With Depth-Normal ConstraintsYan Xu, Xinge Zhu, Jianping Shi, Guofeng Zhang et al.ICCV 2019 · 249 citations
Related papers
- Transfusion: A Novel SLAM Method Focused on Transparent ObjectsYifan Zhu, Jiaxiong Qiu, Bo RenICCV 2021 · 20 citations
- Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface DetectionJiaying Lin, Yuen Hei Yeung, Shuquan Ye, Rynson W. H. LauAAAI 2025 · 15 citations
- TMO: Textured Mesh Acquisition of Objects with a Mobile Device by using Differentiable RenderingJaehoon Choi, Dongki Jung, Taejae Lee, Sangwook Kim et al.CVPR 2023
- Seeing Through the Glass: Neural 3D Reconstruction of Object Inside a Transparent ContainerJinguang Tong, Sundaram Muthu, Fahira Afzal Maken, Chuong Nguyen et al.CVPR 2023
- KeyPose: Multi-View 3D Labeling and Keypoint Estimation for Transparent ObjectsXingyu Liu, Rico Jonschkowski, Anelia Angelova, Kurt KonoligeCVPR 2020
