Scale-Invariant Monocular Depth Estimation via SSI Depth
S. Mahdi H. Miangoleh, Mahesh Kumar Krishna Reddy, Yagiz Aksoy
摘要
Existing methods for scale-invariant monocular depth estimation (SI MDE) often struggle due to the complexity of the task, and limited and non-diverse datasets, hindering generalizability in real-world scenarios. This is while shift-and-scale-invariant (SSI) depth estimation, simplifying the task and enabling training with abundant stereo datasets achieves high performance. We present a novel approach that leverages SSI inputs to enhance SI depth estimation, streamlining the network’s role and facilitating in-the-wild generalization for SI depth estimation while only using a synthetic dataset for training. Emphasizing the generation of high-resolution details, we introduce a novel sparse ordinal loss that substantially improves detail generation in SSI MDE, addressing critical limitations in existing approaches. Through in-the-wild qualitative examples and zero-shot evaluation we substantiate the practical utility of our approach in computational photography applications, showcasing its ability to generate highly detailed SI depth maps and achieve generalization in diverse scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Jasmine: Harnessing Diffusion Prior for Self-supervised Depth EstimationJiyuan Wang, Chunyu Lin, Cheng Guan, Lang Nie 等NeurIPS 2025 · 被引用 26 次
- RPG360: Robust 360 Depth Estimation with Perspective Foundation Models and Graph OptimizationDongki Jung, Jaehoon Choi, Yonghan Lee, Dinesh ManochaNeurIPS 2025 · 被引用 4 次
- StableDepth: Scene-Consistent and Scale-Invariant Monocular DepthZheng Zhang, Lihe Yang, Tianyu Yang, Chaohui Yu 等ICCV 2025 · 被引用 1 次
- Vid2Sim: Realistic and Interactive Simulation from Video for Urban NavigationZiyang Xie, Zhizheng Liu, Zhenghao Peng, Wayne Wu 等CVPR 2025
它引用的顶会 Paper14
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu 等CVPR 2024 · 被引用 847 次
- Hypersim: A Photorealistic Synthetic Dataset for Holistic Indoor Scene UnderstandingMike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar 等ICCV 2021 · 被引用 633 次
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 被引用 487 次
- Omnidata: A Scalable Pipeline for Making Multi-Task Mid-Level Vision Datasets from 3D ScansAinaz Eftekhar, Alexander Sax, Jitendra Malik, Amir ZamirICCV 2021 · 被引用 422 次
相关 Paper
- Hierarchical Normalization for Robust Monocular Depth EstimationChi Zhang, Wei Yin, Billzb Wang, Gang Yu 等NeurIPS 2022 · 被引用 73 次
- BetterDepth: Plug-and-Play Diffusion Refiner for Zero-Shot Monocular Depth EstimationXiang Zhang, Bingxin Ke, Hayko Riemenschneider, Nando Metzger 等NeurIPS 2024 · 被引用 32 次
- Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective SurfacesWonhyeok Choi, Kyumin Hwang, Minwoo Choi, Kiljoon Han 等AAAI 2025 · 被引用 3 次
- UniDepth: Universal Monocular Metric Depth EstimationLuigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segù 等CVPR 2024 · 被引用 122 次
- Robust Geometry-Preserving Depth Estimation Using Differentiable RenderingChi Zhang, Wei Yin, Gang Yu, Zhibin Wang 等ICCV 2023 · 被引用 7 次
