RGB-Depth Fusion GAN for Indoor Depth Completion
Haowen Wang, Mingyuan Wang, Zhengping Che, Zhiyuan Xu, Xiuquan Qiao, Mengshi Qi, Feifei Feng, Jian Tang
Abstract
The raw depth image captured by the indoor depth sen-sor usually has an extensive range of missing depth values due to inherent limitations such as the inability to perceive transparent objects and limited distance range. The incomplete depth map burdens many downstream vision tasks, and a rising number of depth completion methods have been proposed to alleviate this issue. While most existing meth-ods can generate accurate dense depth maps from sparse and uniformly sampled depth maps, they are not suitable for complementing the large contiguous regions of missing depth values, which is common and critical. In this paper, we design a novel two-branch end-to-end fusion network, which takes a pair of RGB and incomplete depth images as input to predict a dense and completed depth map. The first branch employs an encoder-decoder structure to regress the local dense depth values from the raw depth map, with the help of local guidance information extracted from the RGB image. In the other branch, we propose an RGB-depth fusion GAN to transfer the RGB image to the fine-grained textured depth map. We adopt adaptive fusion modules named W-AdaIN to propagate the features across the two branches, and we append a confidence fusion head to fuse the two out-puts of the branches for the final depth map. Extensive ex-periments on NYU-Depth V2 and SUN RGB-D demonstrate that our proposed method clearly improves the depth completion performance, especially in a more realistic setting of indoor environments with the help of the pseudo depth map.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8b2ab635-8cc1-42cf-8e39-4ff2af27d78fCited by top-tier papers9
- Aggregating Feature Point Cloud for Depth CompletionZhu Yu, Zehua Sheng, Zili Zhou, Lun Luo et al.ICCV 2023 · 42 citations
- AGG-Net: Attention Guided Gated-convolutional Network for Depth Image CompletionDongyue Chen, Tingxuan Huang, Zhimin Song, Shizhuo Deng et al.ICCV 2023 · 16 citations
- Learning Long-range Information with Dual-Scale Transformers for Indoor Scene CompletionZiqi Wang, Fei Luo, Xiaoxiao Long, Wenxiao Zhang et al.ICCV 2023 · 6 citations
- DTF-Net: Category-Level Pose Estimation and Shape Reconstruction via Deformable Template FieldHaowen Wang, Zhipeng Fan, Zhen Zhao, Zhengping Che et al.ACM MM 2023 · 6 citations
- Indoor Depth Recovery Based on Deep Unfolding with Non-Local PriorYuhui Dai, Junkang Zhang, Faming Fang, Guixu ZhangICCV 2023 · 2 citations
Builds on2
Related papers
- Improving Depth Completion via Depth Feature UpsamplingYufei Wang, Ge Zhang, Shaoqian Wang, Bo Li et al.CVPR 2024 · 15 citations
- FCFR-Net: Feature Fusion based Coarse-to-Fine Residual Learning for Depth CompletionLina Liu, Xibin Song, Xiaoyang Lyu, Junwei Diao et al.AAAI 2021 · 125 citations
- Depth Completion With Twin Surface Extrapolation at Occlusion BoundariesSaif Muhammad Imran, Xiaoming Liu, Daniel MorrisCVPR 2021
- Robust Multimodal Depth Estimation using Transformer based Generative Adversarial NetworksMd Fahim Faysal Khan, Anusha Devulapally, Siddharth Advani, Vijaykrishnan NarayananACM MM 2022 · 6 citations
- Learning Joint 2D-3D Representations for Depth CompletionYun Chen, Bin Yang, Ming Liang, Raquel UrtasunICCV 2019 · 190 citations
