Distilling Monocular Foundation Model for Fine-grained Depth Completion
Yingping Liang, Yutao Hu, Wenqi Shao, Ying Fu
Abstract
Depth completion involves predicting dense depth maps from sparse LiDAR inputs. However, sparse depth annotations from sensors limit the availability of dense supervision, which is necessary for learning detailed geometric features. In this paper, we propose a two-stage knowledge distillation framework that leverages powerful monocular foundation models to provide dense supervision for depth completion. In the first stage, we introduce a pre-training strategy that generates diverse training data from natural images, which distills geometric knowledge to depth completion. Specifically, we simulate LiDAR scans by utilizing monocular depth and mesh reconstruction, thereby creating training data without requiring ground-truth depth. Besides, monocular depth estimation suffers from inherent scale ambiguity in real-world settings. To address this, in the second stage, we employ a scale- and shift-invariant loss (SSI Loss) to learn real-world scales when fine-tuning on real-world datasets. Our two-stage distillation framework enables depth completion models to harness the strengths of monocular foundation models. Experimental results demonstrate that models trained with our two-stage distillation framework achieve state-of-the-art performance, ranking first place on the KITTI benchmark. Code is available at https://github.com/Sharpiless/DMD3C
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 62a54400-3bc1-4185-b67c-e88270a25582Cited by top-tier papers7
- Event-Driven Dynamic Scene Depth CompletionZhiqiang Yan, Jianhao Jiao, Zhengxue Wang, Gim Hee LeeNeurIPS 2025 · 12 citations
- DuCos: Duality Constrained Depth Super-Resolution via Foundation ModelZhiqiang Yan, Zhengxue Wang, Haoye Dong, Jun Li et al.ICCV 2025 · 3 citations
- The Midas Touch for Metric DepthYu Ma, Zizhan Guo, Zuyi Xiong, Haoran Zhang et al.CVPR 2026 · 2 citations
- PacGDC: Label-Efficient Generalizable Depth Completion with Projection Ambiguity and ConsistencyHaotian Wang, Aoran Xiao, Xiaoqin Zhang, Meng Yang et al.ICCV 2025 · 1 citation
- CARD: A Multi-Modal Automotive Dataset for Dense 3D Reconstruction in Challenging Road TopographyGasser Elazab, Frank Neuhaus, Tilman Koß, Malte Splietker et al.CVPR 2026 · 1 citation
Builds on29
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao et al.NeurIPS 2024 · 2,305 citations
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 1,248 citations
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu et al.CVPR 2024 · 847 citations
- Metric3D: Towards Zero-shot Metric 3D Prediction from A Single ImageWei Yin, Chi Zhang, Hao Chen, Zhipeng Cai et al.ICCV 2023 · 388 citations
- CSPN++: Learning Context and Resource Aware Convolutional Spatial Propagation Networks for Depth CompletionXinjing Cheng, Peng Wang, Chenye Guan, Ruigang YangAAAI 2020 · 270 citations
Related papers
- Weakly Supervised Monocular 3D Detection with a Single-View ImageXueying Jiang, Sheng Jin, Lewei Lu, Xiaoqin Zhang et al.CVPR 2024
- Attention-Based Depth Distillation with 3D-Aware Positional Encoding for Monocular 3D Object DetectionZizhang Wu, Yunzhe Wu, Jian Pu, Xianzhi Li et al.AAAI 2023 · 29 citations
- DesNet: Decomposed Scale-Consistent Network for Unsupervised Depth CompletionZhiqiang Yan, Kun Wang, Xiang Li, Zhenyu Zhang et al.AAAI 2023 · 46 citations
- Exploiting Pseudo Labels in a Self-Supervised Learning Framework for Improved Monocular Depth EstimationAndra Petrovai, Sergiu NedevschiCVPR 2022 · 56 citations
- Distilling Diffusion Models to Efficient 3D LiDAR Scene CompletionShengyuan Zhang, An Zhao, Ling Yang, Zejian Li et al.ICCV 2025 · 1 citation
