GuideFormer: Transformers for Image Guided Depth Completion
Kyeongha Rho, Jinsung Ha, Youngjung Kim
Abstract
Depth completion has been widely studied to predict a dense depth image from its sparse measurement and a single color image. However, most state-of-the-art methods rely on static convolutional neural networks (CNNs) which are not flexible enough for capturing the dynamic nature of input contexts. In this paper, we propose GuideFormer, a fully transformer-based architecture for dense depth completion. We first process sparse depth and color guidance images with separate transformer branches to extract hierarchical and complementary token representations. Each branch consists of a stack of self-attention blocks and has key design features to make our model suitable for the task. We also devise an effective token fusion method based on guided-attention mechanism. It explicitly models information flow between the two branches and captures inter-modal dependencies that cannot be obtained from depth or color image alone. These properties allow GuideFormer to enjoy various visual dependencies and recover precise depth values while preserving fine details. We evaluate GuideFormer on the KITTI dataset containing realworld driving scenes and provide extensive ablation studies. Experimental results demonstrate that our approach significantly outperforms the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8d93ddaf-9e92-4731-ae7e-79be3e2402e8Cited by top-tier papers22
- LRRU: Long-short Range Recurrent Updating Networks for Depth CompletionYufei Wang, Bo Li, Ge Zhang, Qi Liu et al.ICCV 2023 · 89 citations
- DesNet: Decomposed Scale-Consistent Network for Unsupervised Depth CompletionZhiqiang Yan, Kun Wang, Xiang Li, Zhenyu Zhang et al.AAAI 2023 · 46 citations
- Aggregating Feature Point Cloud for Depth CompletionZhu Yu, Zehua Sheng, Zili Zhou, Lun Luo et al.ICCV 2023 · 42 citations
- Tri-Perspective view Decomposition for Geometry-Aware Depth CompletionZhiqiang Yan, Yuankai Lin, Kun Wang, Yupeng Zheng et al.CVPR 2024 · 33 citations
- Bilateral Propagation Network for Depth CompletionJie Tang, Fei-Peng Tian, Boshi An, Jian Li et al.CVPR 2024 · 33 citations
Builds on13
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
Related papers
- CompletionFormer: Depth Completion with Convolutions and Vision TransformersYoumin Zhang, Xianda Guo, Matteo Poggi, Zheng Zhu et al.CVPR 2023
- VoxFormer: Sparse Voxel Transformer for Camera-Based 3D Semantic Scene CompletionYiming Li, Zhiding Yu, Christopher B. Choy, Chaowei Xiao et al.CVPR 2023
- FCFR-Net: Feature Fusion based Coarse-to-Fine Residual Learning for Depth CompletionLina Liu, Xibin Song, Xiaoyang Lyu, Junwei Diao et al.AAAI 2021 · 125 citations
- Robust Multimodal Depth Estimation using Transformer based Generative Adversarial NetworksMd Fahim Faysal Khan, Anusha Devulapally, Siddharth Advani, Vijaykrishnan NarayananACM MM 2022 · 6 citations
- Multi-Frame Self-Supervised Depth with TransformersVitor Guizilini, Rares Ambrus, Dian Chen, Sergey Zakharov et al.CVPR 2022 · 95 citations
