Single Image Depth Prediction Made Better: A Multivariate Gaussian Take
Ce Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte, Luc Van Gool
摘要
Neural-network-based single image depth prediction (SIDP) is a challenging task where the goal is to predict the scene's per-pixel depth at test time. Since the problem, by definition, is ill-posed, the fundamental goal is to come up with an approach that can reliably model the scene depth from a set of training examples. In the pursuit of perfect depth estimation, most existing state-of-the-art learning techniques predict a single scalar depth value per-pixel. Yet, it is well-known that the trained model has accuracy limits and can predict imprecise depth. Therefore, an SIDP approach must be mindful of the expected depth variations in the model's prediction at test time. Accordingly, we introduce an approach that performs continuous modeling of per-pixel depth, where we can predict and reason about the per-pixel depth and its distribution. To this end, we model per-pixel scene depth using a multivariate Gaussian distribution. Moreover, contrary to the existing uncertainty modeling methods-in the same spirit, where per-pixel depth is assumed to be independent, we introduce per-pixel covariance modeling that encodes its depth dependency w.r.t. all the scene points. Unfortunately, per-pixel depth covariance modeling leads to a computationally expensive continuous loss function, which we solve efficiently using the learned low-rank approximation of the overall covariance matrix. Notably, when tested on benchmark datasets such as KITTI, NYU, and SUN-RGB-D, the SIDP model obtained by optimizing our loss function shows state-of-the-art results. Our method's accuracy (named MG) is among the top on the KITTI depth-prediction benchmark leaderboard 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Tri-Perspective view Decomposition for Geometry-Aware Depth CompletionZhiqiang Yan, Yuankai Lin, Kun Wang, Yupeng Zheng 等CVPR 2024 · 被引用 33 次
- DCDepth: Progressive Monocular Depth Estimation in Discrete Cosine DomainKun Wang, Zhiqiang Yan, Junkai Fan, Wanlu Zhu 等NeurIPS 2024 · 被引用 29 次
- A Simple yet Universal Framework for Depth CompletionJin-Hwi Park, Hae-Gon JeonNeurIPS 2024 · 被引用 17 次
- Stereo Risk: A Continuous Modeling Approach to Stereo MatchingCe Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte 等ICML 2024 · 被引用 8 次
- V2Depth: Monocular Depth Estimation via Feature-Level Virtual-View Simulation and RefinementZizhang Wu, Zhuozheng Li, Zhi-Gang Fan, Yunzhe Wu 等ACM MM 2023 · 被引用 4 次
它引用的顶会 Paper19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Deep Evidential RegressionAlexander Amini, Wilko Schwarting, Ava Soleimany, Daniela RusNeurIPS 2020 · 被引用 777 次
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 被引用 487 次
- Neural Window Fully-connected CRFs for Monocular Depth EstimationWeihao Yuan, Xiaodong Gu, Zuozhuo Dai, Siyu Zhu 等CVPR 2022 · 被引用 320 次
相关 Paper
- VA-DepthNet: A Variational Approach to Single Image Depth PredictionCe Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte 等ICLR 2023 · 被引用 17 次
- On the Uncertainty of Self-Supervised Monocular Depth EstimationMatteo Poggi, Filippo Aleotti, Fabio Tosi, Stefano MattocciaCVPR 2020
- Pixelwise Adaptive Discretization with Uncertainty Sampling for Depth CompletionRui Peng, Tao Zhang, Bing Li, Yitong WangACM MM 2022 · 被引用 5 次
- Categorical Depth Distribution Network for Monocular 3D Object DetectionCody Reading, Ali Harakeh, Julia Chae, Steven L. WaslanderCVPR 2021
- P3Depth: Monocular Depth Estimation with a Piecewise Planarity PriorVaishakh Patil, Christos Sakaridis, Alexander Liniger, Luc Van GoolCVPR 2022 · 被引用 144 次
