Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion
Massimiliano Viola, Kevin Qu, Nando Metzger, Bingxin Ke, Alexander Becker, Konrad Schindler, Anton Obukhov
Abstract
Depth completion upgrades sparse depth measurements into dense depth maps guided by a conventional image. Existing methods for this highly ill-posed task operate in tightly constrained settings and tend to struggle when applied to images outside the training domain or when the available depth measurements are sparse, irregularly distributed, or of varying density. Inspired by recent advances in monocular depth estimation, we reframe depth completion as an image-conditional depth map generation guided by sparse measurements. Our method, Marigold-DC, builds on a pretrained latent diffusion model for monocular depth estimation and injects the depth observations as test-time guidance via an optimization scheme that runs in tandem with the iterative inference of denoising diffusion. The method exhibits excellent zero-shot generalization across a diverse range of environments and handles even extremely sparse guidance effectively. Our results suggest that contemporary monocular depth priors greatly robustify depth completion: it may be better to view the task as recovering dense depth from (dense) image pixels, guided by sparse depth; rather than as inpainting (sparse) depth, guided by an image. Project website: https://MarigoldDepthCompletion.github.io/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext accbd697-063e-4345-a0ba-1a2a5fe4a724Cited by top-tier papers17
- Depth Anything with Any PriorZehan Wang, Siyu Chen, Lihe Yang, Jialei Wang et al.ICLR 2026 · 47 citations
- InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit FieldsHao Yu, Haotong Lin, Jiawei Wang, Jiaxin Li et al.CVPR 2026 · 19 citations
- Event-Driven Dynamic Scene Depth CompletionZhiqiang Yan, Jianhao Jiao, Zhengxue Wang, Gim Hee LeeNeurIPS 2025 · 12 citations
- Large Depth Completion Model from Sparse ObservationsZhu Yu, zhengyi zhao, Runmin Zhang, Lingteng Qiu et al.ICLR 2026 · 8 citations
- Orchid: Image Latent Diffusion for Joint Appearance and Geometry GenerationAkshay Krishnan, Xinchen Yan, Vincent Casser, Abhijit KunduICCV 2025 · 8 citations
Builds on40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Repurposing Diffusion-Based Image Generators for Monocular Depth EstimationBingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger et al.CVPR 2024
- Repurposing Marigold for Zero-Shot Metric Depth Estimation via Defocus Blur CuesChinmay Talegaonkar, Nikhil Gandudi Suresh, Zachary Novack, Yash Belhe et al.NeurIPS 2025 · 4 citations
- Zero-shot Depth Completion via Test-time Alignment with Affine-invariant Depth PriorLee Hyoseok, Kyeong Seon Kim, Byung-Ki Kwon, Tae-Hyun OhAAAI 2025 · 11 citations
- OMNI-DC: Highly Robust Depth Completion with Multiresolution Depth IntegrationYiming Zuo, Willow Yang, Zeyu Ma, Jia DengICCV 2025 · 5 citations
- Iris: Integrating Language into Diffusion-based Monocular Depth EstimationZiyao Zeng, Jingcheng Ni, Daniel Wang, Patrick Rim et al.CVPR 2026
