AggNet for Self-supervised Monocular Depth Estimation: Go An Aggressive Step Furthe
Zhi Chen, Xiaoqing Ye, Liang Du, Wei Yang, Liusheng Huang, Xiao Tan, Zhenbo Shi, Fumin Shen, Errui Ding
Abstract
Without appealing to exhaustive labeled data, self-supervised monocular depth estimation (MDE) plays a fundamental role in computer vision. Previous methods usually adopt a one-stage MDE network, which is insufficient to achieve high performance. In this paper, we dig deep into this task to propose an aggressive framework termed AggNet. The framework is based on a training-only progressive two-stage module to perform pseudo counter-surveillance as well as a simple yet effective dual-warp loss function between image pairs. In particular, we first propose a residual module, which follows the MDE network to learn a refined depth. The residual module takes both the initial depth generated from MDE and the initial color image as input to generate refined depth with residual depth learning. Then, the refined depth is leveraged to supervise the initial depth simultaneously during the training period. For inference, only the MDE network is retained to regress depth from a single image, which gains better performance without introducing extra computation. In addition to self-distillation loss, a simple yet effective dual-warp consistency loss is introduced to encourage the MDE network to keep depth consistency between stereo image pairs. Extensive experiments show that our AggNet achieves state-of-the-art performance on the KITTI and Make3D datasets.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 25d4f232-a34b-4d62-9cc1-f30d73dff5a1Cited by top-tier papers1
Ask how each one uses itRelated papers
- Revealing the Reciprocal Relations between Self-Supervised Stereo and Monocular Depth EstimationZhi Chen, Xiaoqing Ye, Wei Yang, Zhenbo Xu et al.ICCV 2021 · 34 citations
- Exploiting Pseudo Labels in a Self-Supervised Learning Framework for Improved Monocular Depth EstimationAndra Petrovai, Sergiu NedevschiCVPR 2022 · 56 citations
- Learning Occlusion-aware Coarse-to-Fine Depth Map for Self-supervised Monocular Depth EstimationZhengming Zhou, Qiulei DongACM MM 2022 · 21 citations
- PLADE-Net: Towards Pixel-Level Accuracy for Self-Supervised Single-View Depth Estimation With Neural Positional Encoding and Distilled Matting LossJuan Luis Gonzalez, Munchurl KimCVPR 2021
- DualNet: Robust Self-Supervised Stereo Matching with Pseudo-Label SupervisionYun Wang, Jiahao Zheng, Chenghao Zhang, Zhanjie Zhang et al.AAAI 2025 · 12 citations
