Zero-shot Depth Completion via Test-time Alignment with Affine-invariant Depth Prior
Lee Hyoseok, Kyeong Seon Kim, Byung-Ki Kwon, Tae-Hyun Oh
Abstract
Depth completion, predicting dense depth maps from sparse depth measurements, is an ill-posed problem requiring prior knowledge. Recent methods adopt learning-based approaches to implicitly capture priors, but the priors primarily fit in-domain data and do not generalize well to out-of-domain scenarios. To address this, we propose a zero-shot depth completion method composed of an affine-invariant depth diffusion model and test-time alignment. We use pre-trained depth diffusion models as depth prior knowledge, which implicitly understand how to fill in depth for scenes. Our approach aligns the affine-invariant depth prior with metric-scale sparse measurements, enforcing them as hard constraints via an optimization loop at test-time. Our zero-shot depth completion method demonstrates generalization across various domain datasets, achieving up to a 21% average performance improvement over the previous state-of-the-art methods while enhancing spatial understanding by sharpening scene details. We demonstrate that aligning a monocular affine-invariant depth prior with sparse metric measurements is a sufficient strategy to achieve domain-generalizable depth completion without relying on extensive training datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Marigold-DC: Zero-Shot Monocular Depth Completion with Guided DiffusionMassimiliano Viola, Kevin Qu, Nando Metzger, Bingxin Ke et al.ICCV 2025 · 16 citations
- Test-Time Prompt Tuning for Zero-Shot Depth CompletionChanhwi Jeong, Inhwan Bae, Jin-Hwi Park, Hae-Gon JeonICCV 2025 · 3 citations
- JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion TransformersByung-Ki Kwon, Qi Dai, Lee Hyoseok, Chong Luo et al.ICCV 2025 · 3 citations
Builds on32
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
Related papers
- BetterDepth: Plug-and-Play Diffusion Refiner for Zero-Shot Monocular Depth EstimationXiang Zhang, Bingxin Ke, Hayko Riemenschneider, Nando Metzger et al.NeurIPS 2024 · 32 citations
- Depth Anything with Any PriorZehan Wang, Siyu Chen, Lihe Yang, Jialei Wang et al.ICLR 2026 · 47 citations
- UniDepth: Universal Monocular Metric Depth EstimationLuigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segù et al.CVPR 2024 · 122 citations
- Hierarchical Normalization for Robust Monocular Depth EstimationChi Zhang, Wei Yin, Billzb Wang, Gang Yu et al.NeurIPS 2022 · 73 citations
- OMNI-DC: Highly Robust Depth Completion with Multiresolution Depth IntegrationYiming Zuo, Willow Yang, Zeyu Ma, Jia DengICCV 2025 · 5 citations
