HR-Depth: High Resolution Self-Supervised Monocular Depth Estimation
Xiaoyang Lyu, Liang Liu, Mengmeng Wang, Xin Kong, Lina Liu, Yong Liu, Xinxin Chen, Yi Yuan
Abstract
Self-supervised learning shows great potential in monocular depth estimation, using image sequences as the only source of supervision. Although people try to use the high-resolution image for depth estimation, the accuracy of prediction has not been significantly improved. In this work, we find the core reason comes from the inaccurate depth estimation in large gradient regions, making the bilinear interpolation error gradually disappear as the resolution increases. To obtain more accurate depth estimation in large gradient regions, it is necessary to obtain high-resolution features with spatial and semantic information. Therefore, we present an improved DepthNet, HR-Depth, with two effective strategies: (1) redesign the skip-connection in DepthNet to get better highresolution features and (2) propose feature fusion Squeezeand-Excitation(fSE) module to fuse feature more efficiently. Using Resnet-18 as the encoder, HR-Depth surpasses all previous state-of-the-art(SoTA) methods with the least parameters at both high and low resolution. Moreover, previous state-of-the-art methods are based on fairly complex and deep networks with a mass of parameters which limits their real applications. Thus we also construct a lightweight network which uses MobileNetV3 as encoder. Experiments show that the lightweight network can perform on par with many large models like Monodepth2 at high-resolution with only 20% parameters. All codes and models will be available at https: //github.com/shawLyu/HR-Depth .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b4dbbece-47ec-4c20-afbd-2d8455fae5a5Cited by top-tier papers33
- Fine-grained Semantics-aware Representation Enhancement for Self-supervised Monocular Depth EstimationHyunyoung Jung, Eunhyeok Park, Sungjoo YooICCV 2021 · 133 citations
- Deep Digging into the Generalization of Self-Supervised Monocular Depth EstimationJinwoo Bae, Sungho Moon, Sunghoon ImAAAI 2023 · 127 citations
- Self-supervised Monocular Depth Estimation for All Day Images using Domain SeparationLina Liu, Xibin Song, Mengmeng Wang, Yong Liu et al.ICCV 2021 · 95 citations
- Excavating the Potential Capacity of Self-Supervised Monocular Depth EstimationRui Peng, Ronggang Wang, Yawen Lai, Luyang Tang et al.ICCV 2021 · 93 citations
- SimIPU: Simple 2D Image and 3D Point Cloud Unsupervised Pre-training for Spatial-Aware Visual RepresentationsZhenyu Li, Zehui Chen, Ang Li, Liangji Fang et al.AAAI 2022 · 78 citations
Builds on4
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Semantically-Guided Representation Learning for Self-Supervised Monocular DepthVitor Guizilini, Rui Hou, Jie Li, Rares Ambrus et al.ICLR 2020 · 264 citations
- Unsupervised High-Resolution Depth Learning From Videos With Dual NetworksJunsheng Zhou, Yuwang Wang, Kaihuai Qin, Wenjun ZengICCV 2019 · 77 citations
Related papers
- R-MSFM: Recurrent Multi-Scale Feature Modulation for Monocular Depth EstimatingZhongkai Zhou, Xinnan Fan, Pengfei Shi, Yuanxue XinICCV 2021 · 150 citations
- Multi-Resolution Monocular Depth Map Fusion by Self-Supervised Gradient-Based CompositionYaqiao Dai, Renjiao Yi, Chenyang Zhu, Hongjun He et al.AAAI 2023 · 8 citations
- Lite-Mono: A Lightweight CNN and Transformer Architecture for Self-Supervised Monocular Depth EstimationNing Zhang, Francesco Nex, George Vosselman, Norman KerleCVPR 2023
- Self-Supervised Monocular Trained Depth Estimation Using Self-Attention and Discrete Disparity VolumeAdrian Johnston, Gustavo CarneiroCVPR 2020
- Two-in-One Depth: Bridging the Gap Between Monocular and Binocular Self-supervised Depth EstimationZhengming Zhou, Qiulei DongICCV 2023 · 16 citations
