VA-DepthNet: A Variational Approach to Single Image Depth Prediction
Ce Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte, Luc Van Gool
摘要
We introduce VA-DepthNet, a simple, effective, and accurate deep neural network approach for the single-image depth prediction (SIDP) problem. The proposed approach advocates using classical first-order variational constraints for this problem. While state-of-the-art deep neural network methods for SIDP learn the scene depth from images in a supervised setting, they often overlook the invaluable invariances and priors in the rigid scene space, such as the regularity of the scene. The paper's main contribution is to reveal the benefit of classical and well-founded variational constraints in the neural network design for the SIDP task. It is shown that imposing first-order variational constraints in the scene space together with popular encoder-decoder-based network architecture design provides excellent results for the supervised SIDP task. The imposed first-order variational constraint makes the network aware of the depth gradient in the scene space, i.e., regularity. The paper demonstrates the usefulness of the proposed approach via extensive evaluation and ablation analysis over several benchmark datasets, such as KITTI, NYU Depth V2, and SUN RGB-D. The VA-DepthNet at test time shows considerable improvements in depth prediction accuracy compared to the prior art and is accurate also at high-frequency regions in the scene space. At the time of writing this paper, our method-labeled as VA-DepthNet, when tested on the KITTI depth-prediction evaluation set benchmarks, shows state-of-the-art results, and is the top-performing published approach 1 2 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Towards Zero-Shot Scale-Aware Monocular Depth EstimationVitor Guizilini, Igor Vasiljevic, Dian Chen, Rares Ambrus 等ICCV 2023 · 被引用 129 次
- DCDepth: Progressive Monocular Depth Estimation in Discrete Cosine DomainKun Wang, Zhiqiang Yan, Junkai Fan, Wanlu Zhu 等NeurIPS 2024 · 被引用 29 次
- UC-NERF: Neural Radiance Field for Under-Calibrated Multi-View Cameras in Autonomous DrivingKai Cheng, Xiaoxiao Long, Wei Yin, Jin Wang 等ICLR 2024 · 被引用 26 次
- RSA: Resolving Scale Ambiguities in Monocular Depth Estimators through Language DescriptionsZiyao Zeng, Yangchao Wu, Hyoungseob Park, Daniel Wang 等NeurIPS 2024 · 被引用 26 次
- Rethinking Multi-Domain Generalization with A General Learning ObjectiveZhaorui Tan, Xi Yang, Kaizhu HuangCVPR 2024 · 被引用 22 次
它引用的顶会 Paper19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
相关 Paper
- Single Image Depth Prediction Made Better: A Multivariate Gaussian TakeCe Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte 等CVPR 2023
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 被引用 487 次
- How Do Neural Networks See Depth in Single Images?Tom van Dijk, Guido de CroonICCV 2019 · 被引用 210 次
- Learning Occlusion-aware Coarse-to-Fine Depth Map for Self-supervised Monocular Depth EstimationZhengming Zhou, Qiulei DongACM MM 2022 · 被引用 21 次
- V2Depth: Monocular Depth Estimation via Feature-Level Virtual-View Simulation and RefinementZizhang Wu, Zhuozheng Li, Zhi-Gang Fan, Yunzhe Wu 等ACM MM 2023 · 被引用 4 次
