Generating and Exploiting Probabilistic Monocular Depth Estimates
Zhihao Xia, Patrick Sullivan, Ayan Chakrabarti
Abstract
Beyond depth estimation from a single image, the monocular cue is useful in a broader range of depth inference applications and settings-such as when one can leverage other available depth cues for improved accuracy. Currently, different applications, with different inference tasks and combinations of depth cues, are solved via different specialized networks-trained separately for each application. Instead, we propose a versatile task-agnostic monocular model that outputs a probability distribution over scene depth given an input color image, as a sample approximation of outputs from a patch-wise conditional VAE. We show that this distributional output can be used to enable a variety of inference tasks in different settings, without needing to retrain for each application. Across a diverse set of applications (depth completion, user guided estimation, etc.), our common model yields results with high accuracycomparable to or surpassing that of state-of-the-art methods dependent on application-specific networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- Transformer-Based Attention Networks for Continuous Pixel-Wise PredictionGuanglei Yang, Hao Tang, Mingli Ding, Nicu Sebe et al.ICCV 2021 · 246 citations
- WorDepth: Variational Language Prior for Monocular Depth EstimationZiyao Zeng, Daniel Wang, Fengyu Yang, Hyoungseob Park et al.CVPR 2024 · 20 citations
- Adaptive Gating for Single-Photon 3D ImagingRyan Po, Adithya Pediredla, Ioannis GkioulekasCVPR 2022 · 16 citations
- Learning Structured Gaussians to Approximate Deep EnsemblesIvor J. A. Simpson, Sara Vicente, Neill D. F. CampbellCVPR 2022 · 8 citations
- A Dark Flash Normal CameraZhihao Xia, Jason Lawrence, Supreeth AcharICCV 2021 · 6 citations
Related papers
- UniDepth: Universal Monocular Metric Depth EstimationLuigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segù et al.CVPR 2024 · 122 citations
- Marigold-DC: Zero-Shot Monocular Depth Completion with Guided DiffusionMassimiliano Viola, Kevin Qu, Nando Metzger, Bingxin Ke et al.ICCV 2025 · 16 citations
- Learning a Depth Covariance FunctionEric Dexheimer, Andrew J. DavisonCVPR 2023
- Orchid: Image Latent Diffusion for Joint Appearance and Geometry GenerationAkshay Krishnan, Xinchen Yan, Vincent Casser, Abhijit KunduICCV 2025 · 8 citations
- Depth Anything with Any PriorZehan Wang, Siyu Chen, Lihe Yang, Jialei Wang et al.ICLR 2026 · 47 citations
