Towards Interpretable Deep Networks for Monocular Depth Estimation
Zunzhi You, Yi-Hsuan Tsai, Wei-Chen Chiu, Guanbin Li
Abstract
Deep networks for Monocular Depth Estimation (MDE) have achieved promising performance recently and it is of great importance to further understand the interpretability of these networks. Existing methods attempt to provide posthoc explanations by investigating visual cues, which may not explore the internal representations learned by deep networks. In this paper, we find that some hidden units of the network are selective to certain ranges of depth, and thus such behavior can be served as a way to interpret the internal representations. Based on our observations, we quantify the interpretability of a deep MDE network by the depth selectivity of its hidden units. Moreover, we then propose a method to train interpretable MDE deep networks without changing their original architectures, by assigning a depth range for each unit to select. Experimental results demonstrate that our method is able to enhance the interpretability of deep MDE networks by largely improving the depth selectivity of their units, while not harming or even improving the depth estimation accuracy. We further provide comprehensive analysis to show the reliability of selective units, the applicability of our method on different layers, models, and datasets, and a demonstration on analysis of model error. Source code and models are available at https:// github.com/youzunzhi/InterpretableMDE .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cb08152c-66a7-4eed-b347-71bf2aae810bCited by top-tier papers3
- CodedEvents: Optimal Point-Spread-Function Engineering for 3D-Tracking with Event CamerasSachin Shah, Matthew A. Chan, Haoming Cai, Jingxi Chen et al.CVPR 2024 · 4 citations
- CMoB: Modality Valuation via Causal Effect for Balanced Multimodal LearningJun Wang, Fuyuan Cao, Zhixin Xue, Xingwang Zhao et al.NeurIPS 2025 · 4 citations
- Personalized Semantics Excitation for Federated Image ClassificationHaifeng Xia, Kai Li, Zhengming DingICCV 2023 · 1 citation
Builds on5
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 487 citations
- How Do Neural Networks See Depth in Single Images?Tom van Dijk, Guido de CroonICCV 2019 · 210 citations
- Visualization of Convolutional Neural Networks for Monocular Depth EstimationJunjie Hu, Yan Zhang, Takayuki OkataniICCV 2019 · 91 citations
- Towards Interpretable Object Detection by Unfolding Latent StructuresTianfu Wu, Xi SongICCV 2019 · 28 citations
- The Edge of Depth: Explicit Constraints Between Segmentation and DepthShengjie Zhu, Garrick Brazil, Xiaoming LiuCVPR 2020
Related papers
- DepthCues: Evaluating Monocular Depth Perception in Large Vision ModelsDuolikun Danier, Mehmet Aygün, Changjian Li, Hakan Bilen et al.CVPR 2025
- Measuring Per-Unit Interpretability at Scale Without HumansRoland S. Zimmermann, David A. Klindt, Wieland BrendelNeurIPS 2024 · 5 citations
- Hierarchical Normalization for Robust Monocular Depth EstimationChi Zhang, Wei Yin, Billzb Wang, Gang Yu et al.NeurIPS 2022 · 73 citations
- Learning Regularizer for Monocular Depth Estimation with Adversarial GuidanceGuibao Shen, Yingkui Zhang, Jialu Li, Mingqiang Wei et al.ACM MM 2021 · 7 citations
- Single Image Depth Prediction With Wavelet DecompositionMichaël Ramamonjisoa, Michael Firman, Jamie Watson, Vincent Lepetit et al.CVPR 2021
