3D Distillation: Improving Self-Supervised Monocular Depth Estimation on Reflective Surfaces
Xuepeng Shi, Georgi Dikov, Gerhard Reitmayr, Tae-Kyun Kim, Mohsen Ghafoorian
Abstract
Self-supervised monocular depth estimation (SSMDE) aims at predicting the dense depth maps of monocular images, by learning to minimize a photometric loss using spatially neighboring image pairs during training. While SSMDE offers a significant scalability advantage over supervised approaches, it performs poorly on reflective surfaces as the photometric constancy assumption of the photometric loss is violated. We note that the appearance of reflective surfaces is view-dependent and often there are views of such surfaces in the training data that are not contaminated by strong specular reflections. Thus, reflective surfaces can be accurately reconstructed by aggregating the predicted depth of these views. Motivated by this observation, we propose 3D distillation: a novel training framework that utilizes the projected depth of reconstructed reflective surfaces to generate reasonably accurate depth pseudo-labels. To identify those surfaces automatically, we employ an uncertainty-guided depth fusion method, combining the smoother and more accurate projected depth on reflective surfaces and the detailed predicted depth elsewhere. In our experiments using the ScanNet and 7-Scenes datasets, we show that 3D distillation not only significantly improves the prediction accuracy, especially on the problematic surfaces, but also that it generalizes well over various underlying network architectures and to new datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 953fbb00-cf23-4db8-9b0f-5b0ed2ab20eaCited by top-tier papers4
- Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective SurfacesWonhyeok Choi, Kyumin Hwang, Minwoo Choi, Kiljoon Han et al.AAAI 2025 · 3 citations
- Instance-Level Video Depth in Groups Beyond OcclusionsYuan Liang, Yang Zhou, Ziming Sun, Tianyi Xiang et al.ICCV 2025
- Self-supervised Monocular Depth Estimation Robust to Reflective Surface Leveraged by Triplet MiningWonhyeok Choi, Kyumin Hwang, Wei Peng, Minwoo Choi et al.ICLR 2025
- CLIPDet3D: Vision-Language Collaborative Distillation for 3D Object DetectionJiaqi Zhao, Huanfeng Hu, Yong Zhou, Wen-Liang Du et al.AAAI 2026
Builds on16
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Omnidata: A Scalable Pipeline for Making Multi-Task Mid-Level Vision Datasets from 3D ScansAinaz Eftekhar, Alexander Sax, Jitendra Malik, Amir ZamirICCV 2021 · 422 citations
- HR-Depth: High Resolution Self-Supervised Monocular Depth EstimationXiaoyang Lyu, Liang Liu, Mengmeng Wang, Xin Kong et al.AAAI 2021 · 341 citations
- MPViT: Multi-Path Vision Transformer for Dense PredictionYoungwan Lee, Jonghee Kim, Jeffrey Willette, Sung Ju HwangCVPR 2022 · 339 citations
Related papers
- Exploiting Pseudo Labels in a Self-Supervised Learning Framework for Improved Monocular Depth EstimationAndra Petrovai, Sergiu NedevschiCVPR 2022 · 56 citations
- Weakly Supervised Monocular 3D Detection with a Single-View ImageXueying Jiang, Sheng Jin, Lewei Lu, Xiaoqin Zhang et al.CVPR 2024
- Revealing the Reciprocal Relations between Self-Supervised Stereo and Monocular Depth EstimationZhi Chen, Xiaoqing Ye, Wei Yang, Zhenbo Xu et al.ICCV 2021 · 34 citations
- AggNet for Self-supervised Monocular Depth Estimation: Go An Aggressive Step FurtheZhi Chen, Xiaoqing Ye, Liang Du, Wei Yang et al.ACM MM 2021 · 6 citations
- Toward Practical Monocular Indoor Depth EstimationCho-Ying Wu, Jialiang Wang, Michael Hall, Ulrich Neumann et al.CVPR 2022 · 68 citations
