DeCoTR: Enhancing Depth Completion with 2D and 3D Attentions
Yunxiao Shi, Manish Kumar Singh, Hong Cai, Fatih Porikli
摘要
In this paper, we introduce a novel approach that har-nesses both 2D and 3D attentions to enable highly accurate depth completion without requiring iterative spatial propa-gations. Specifically, we first enhance a baseline convolutional depth completion model by applying attention to 2D features in the bottleneck and skip connections. This effectively improves the performance of this simple network and sets it on par with the latest, complex transformer-based models. Leveraging the initial depths and features from this network, we uplift the 2D features to form a 3D point cloud and construct a 3D point transformer to process it, allowing the model to explicitly learn and exploit 3D geometric features. In addition, we propose normalization techniques to process the point cloud, which improves learning and leads to better accuracy than directly using point transformers off the shelf. Furthermore, we incorporate global attention on downsampled point cloud features, which enables long-range context while still being computationally feasible. We evaluate our method, DeCoTr, on established depth Completion benchmarks, including NYU Depth V2 and KITTI, showcasing that it sets new state-of-the-art performance. We further conduct zero-shot evaluations on ScanNet and DDAD benchmarks and demonstrate that DeCoTR has su-perior generalizability compared to existing approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper16
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu 等NeurIPS 2022 · 被引用 924 次
- Depth Completion From Sparse LiDAR Data With Depth-Normal ConstraintsYan Xu, Xinge Zhu, Jianping Shi, Guofeng Zhang 等ICCV 2019 · 被引用 249 次
- Learning Joint 2D-3D Representations for Depth CompletionYun Chen, Bin Yang, Ming Liang, Raquel UrtasunICCV 2019 · 被引用 190 次
- Dynamic Spatial Propagation Network for Depth CompletionYuankai Lin, Tao Cheng, Qi Zhong, Wending Zhou 等AAAI 2022 · 被引用 155 次
相关 Paper
- CompletionFormer: Depth Completion with Convolutions and Vision TransformersYoumin Zhang, Xianda Guo, Matteo Poggi, Zheng Zhu 等CVPR 2023
- Aggregating Feature Point Cloud for Depth CompletionZhu Yu, Zehua Sheng, Zili Zhou, Lun Luo 等ICCV 2023 · 被引用 42 次
- Context and Geometry Aware Voxel Transformer for Semantic Scene CompletionZhu Yu, Runmin Zhang, Jiacheng Ying, Junchen Yu 等NeurIPS 2024 · 被引用 73 次
- GeoFormer: Learning Point Cloud Completion with Tri-Plane Integrated TransformerJinpeng Yu, Binbin Huang, Yuxuan Zhang, Huaxia Li 等ACM MM 2024 · 被引用 14 次
- PointCFormer: A Relation-Based Progressive Feature Extraction Network for Point Cloud CompletionYi Zhong, Weize Quan, Dong-Ming Yan, Jie Jiang 等AAAI 2025 · 被引用 3 次
