Unsupervised Photometric-Consistent Depth Estimation from Endoscopic Monocular Video
Shijie Li, Weijun Lin, Qingyuan Xiang, Yunbin Tu, Shitan Asu, Zheng Li
Abstract
Recent advancements in unsupervised monocular depth estimation typically rely on an assumption that image photometry remains consistent across consecutive frames. However, this assumption often fails in endoscopic scenes due to: 1) local photometric inconsistency caused by specular reflections creating highlights; and 2) global photometric inconsistency resulting from the simultaneous movement of the light source and the camera. Since unsupervised depth estimation methods rely on appearance discrepancies between frames as a supervisory signal, these photometric inconsistencies inevitably deteriorate loss function calculation. In this paper, our goal is to obtain a strong and reliable supervisory signal for achieving photometric-consistent depth estimation. To this end, for local photometric inconsistency, we utilize the specular reflection model to introduce a Highlight Loss for handling the estimation of highlight regions. For global photometric inconsistency, we design a Photometric Match module, which utilizes the spotlight illumination model to derive an analytical expression, achieving photometric alignment across different frames. Unlike previous works that introduce additional optical flow or networks, our method is simpler and more efficient. Extensive experiments demonstrate our method achieves the state-of-the-art results on C3VD, SCARED and SERV-CT datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6741d634-1c92-4910-bf6a-0e5216e1d28dCited by top-tier papers1
Ask how each one uses itBuilds on9
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu et al.CVPR 2024 · 847 citations
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 487 citations
- Self-Supervised Learning With Geometric Constraints in Monocular Video: Connecting Flow, Depth, and CameraYuhua Chen, Cordelia Schmid, Cristian SminchisescuICCV 2019 · 265 citations
- ACDNet: Adaptively Combined Dilated Convolution for Monocular Panorama Depth EstimationChuanqing Zhuang, Zhengda Lu, Yiqun Wang, Jun Xiao et al.AAAI 2022 · 73 citations
Related papers
- DeLightMono: Enhancing Self-Supervised Monocular Depth Estimation in Endoscopy by Decoupling Uneven IlluminationMingyang Ou, Haojin Li, Yifeng Zhang, Ke Niu et al.AAAI 2026
- Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective SurfacesWonhyeok Choi, Kyumin Hwang, Minwoo Choi, Kiljoon Han et al.AAAI 2025 · 3 citations
- Enhancing Self-supervised Monocular Depth Estimation via Incorporating Robust ConstraintsRui Li, Xiantuo He, Yu Zhu, Xianjun Li et al.ACM MM 2020 · 16 citations
- 3D Distillation: Improving Self-Supervised Monocular Depth Estimation on Reflective SurfacesXuepeng Shi, Georgi Dikov, Gerhard Reitmayr, Tae-Kyun Kim et al.ICCV 2023 · 10 citations
- CL-MVSNet: Unsupervised Multi-view Stereo with Dual-level Contrastive LearningKaiqiang Xiong, Rui Peng, Zhe Zhang, Tianxing Feng et al.ICCV 2023 · 24 citations
