Learning Single Camera Depth Estimation Using Dual-Pixels
Rahul Garg, Neal Wadhwa, Sameer Ansari, Jonathan T. Barron
摘要
Deep learning techniques have enabled rapid progress in monocular depth estimation, but their quality is limited by the ill-posed nature of the problem and the scarcity of high quality datasets. We estimate depth from a single cam-era by leveraging the dual-pixel auto-focus hardware that is increasingly common on modern camera sensors. Classic stereo algorithms and prior learning-based depth estimation techniques underperform when applied on this dual-pixel data, the former due to too-strong assumptions about RGB image matching, and the latter due to not leveraging the understanding of optics of dual-pixel image formation. To allow learning based methods to work well on dual-pixel imagery, we identify an inherent ambiguity in the depth estimated from dual-pixel cues, and develop an approach to estimate depth up to this ambiguity. Using our approach, existing monocular depth estimation techniques can be effectively applied to dual-pixel data, and much smaller models can be constructed that still infer high quality depth. To demonstrate this, we capture a large dataset of in-the-wild 5-viewpoint RGB images paired with corresponding dual-pixel data, and show how view supervision with this data can be used to learn depth up to the unknown ambiguities. On our new task, our model is 30% more accurate than any prior work on learning-based monocular or stereoscopic depth estimation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Single Image Defocus Deblurring Using Kernel-Sharing Parallel Atrous ConvolutionsHyeongseok Son, Junyong Lee, Sunghyun Cho, Seungyong LeeICCV 2021 · 被引用 124 次
- Learning to Reduce Defocus Blur by Realistically Modeling Dual-Pixel DataAbdullah Abuolaim, Mauricio Delbracio, Damien Kelly, Michael S. Brown 等ICCV 2021 · 被引用 72 次
- Gaussian Kernel Mixture Network for Single Image Defocus DeblurringYuhui Quan, Zicong Wu, Hui JiNeurIPS 2021 · 被引用 65 次
- SLIDE: Single Image 3D Photography with Soft Layering and Depth-aware InpaintingVarun Jampani, Huiwen Chang, Kyle Sargent, Abhishek Kar 等ICCV 2021 · 被引用 63 次
- Defocus Map Estimation and Deblurring from a Single Dual-Pixel ImageShumian Xin, Neal Wadhwa, Tianfan Xue, Jonathan T. Barron 等ICCV 2021 · 被引用 47 次
相关 Paper
- Learning to AutofocusCharles Herrmann, Richard Strong Bowen, Neal Wadhwa, Rahul Garg 等CVPR 2020
- Spatio-Focal Bidirectional Disparity Estimation from a Dual-Pixel ImageDonggun Kim, Hyeonjoong Jang, Inchul Kim, Min H. KimCVPR 2023
- DepthInSpace: Exploitation and Fusion of Multiple Video Frames for Structured-Light Depth EstimationMohammad Mahdi Johari, Camilla Carta, François FleuretICCV 2021 · 被引用 12 次
- Exploring Positional Characteristics of Dual-Pixel Data for Camera AutofocusMyungsub Choi, Hana Lee, Hyong-Euk LeeICCV 2023 · 被引用 9 次
- Dual Pixel Exploration: Simultaneous Depth Estimation and Image RestorationLiyuan Pan, Shah Chowdhury, Richard Hartley, Miaomiao Liu 等CVPR 2021
