SDC-Depth: Semantic Divide-and-Conquer Network for Monocular Depth Estimation
Lijun Wang, Jianming Zhang, Oliver Wang, Zhe Lin, Huchuan Lu
Abstract
Monocular depth estimation is an ill-posed problem, and as such critically relies on scene priors and semantics. Due to its complexity, we propose a deep neural network model based on a semantic divide-and-conquer approach. Our model decomposes a scene into semantic segments, such as object instances and background stuff classes, and then predicts a scale and shift invariant depth map for each semantic segment in a canonical space. Semantic segments of the same category share the same depth decoder, so the global depth prediction task is decomposed into a series of category-specific ones, which are simpler to learn and easier to generalize to new scene types. Finally, our model stitches each local depth segment by predicting its scale and shift based on the global context of the image. The model is trained end-to-end using a multi-task loss for panoptic segmentation and depth prediction, and is therefore able to leverage large-scale panoptic segmentation datasets to boost its semantic understanding. We validate the effectiveness of our approach and show state-of-the-art performance on three benchmark datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3bccc4c4-6791-4bf3-8b79-b49290f64911Cited by top-tier papers24
- TransWeather: Transformer-based Restoration of Images Degraded by Adverse Weather ConditionsJeya Maria Jose Valanarasu, Rajeev Yasarla, Vishal M. PatelCVPR 2022 · 350 citations
- SAVi++: Towards End-to-End Object-Centric Learning from Real-World VideosGamaleldin F. Elsayed, Aravindh Mahendran, Sjoerd van Steenkiste, Klaus Greff et al.NeurIPS 2022 · 218 citations
- Self-supervised Monocular Depth Estimation for All Day Images using Domain SeparationLina Liu, Xibin Song, Mengmeng Wang, Yong Liu et al.ICCV 2021 · 95 citations
- All in Tokens: Unifying Output Space of Visual Tasks via Soft TokenJia Ning, Chen Li, Zheng Zhang, Chunyu Wang et al.ICCV 2023 · 64 citations
- MGNet: Monocular Geometric Scene Understanding for Autonomous DrivingMarkus Schön, Michael Buchholz, Klaus DietmayerICCV 2021 · 60 citations
Builds on1
Related papers
- PanopticDepth: A Unified Framework for Depth-aware Panoptic SegmentationNaiyu Gao, Fei He, Jian Jia, Yanhu Shan et al.CVPR 2022 · 27 citations
- Panoptic 3D Scene Reconstruction From a Single RGB ImageManuel Dahnert, Ji Hou, Matthias Nießner, Angela DaiNeurIPS 2021 · 106 citations
- Robust Geometry-Preserving Depth Estimation Using Differentiable RenderingChi Zhang, Wei Yin, Gang Yu, Zhibin Wang et al.ICCV 2023 · 7 citations
- VIP-DeepLab: Learning Visual Perception With Depth-Aware Video Panoptic SegmentationSiyuan Qiao, Yukun Zhu, Hartwig Adam, Alan L. Yuille et al.CVPR 2021
- Semantically-Guided Representation Learning for Self-Supervised Monocular DepthVitor Guizilini, Rui Hou, Jie Li, Rares Ambrus et al.ICLR 2020 · 264 citations
