Deep Optics for Monocular Depth Estimation and 3D Object Detection
Julie Chang, Gordon Wetzstein
摘要
Depth estimation and 3D object detection are critical for scene understanding but remain challenging to perform with a single image due to the loss of 3D information during image capture. Recent models using deep neural networks have improved monocular depth estimation performance, but there is still difficulty in predicting absolute depth and generalizing outside a standard dataset. Here we introduce the paradigm of deep optics, i.e. end-to-end design of optics and image processing, to the monocular depth estimation problem, using coded defocus blur as an additional depth cue to be decoded by a neural network. We evaluate several optical coding strategies along with an end-to-end optimization scheme for depth estimation on three datasets, including NYU Depth v2 and KITTI. We find an optimized freeform lens design yields the best results, but chromatic aberration from a singlet lens offers significantly improved performance as well. We build a physical prototype and validate that chromatic aberrations improve depth estimation on real-world results. In addition, we train object detection networks on the KITTI dataset and show that the lens optimized for depth estimation also results in improved 3D object detection performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper31
- Single-shot Hyperspectral-Depth Imaging with Learned Diffractive OpticsSeung-Hwan Baek, Hayato Ikoma, Daniel S. Jeon, Yuqi Li 等ICCV 2021 · 被引用 109 次
- Learning Privacy-preserving Optics for Human Pose EstimationCarlos Hinojosa, Juan Carlos Niebles, Henry ArguelloICCV 2021 · 被引用 63 次
- Quantization-aware Deep Optics for Diffractive Snapshot Hyperspectral ImagingLingen Li, Lizhi Wang, Weitao Song, Lei Zhang 等CVPR 2022 · 被引用 39 次
- Seeing through obstructions with diffractive cloakingZheng Shi, Yuval Bahat, Seung-Hwan Baek, Qiang Fu 等SIGGRAPH 2022 · 被引用 34 次
- Time-Multiplexed Coded Aperture Imaging: Learned Coded Aperture and Pixel Exposures for Compressive Imaging SystemsEdwin Vargas, Julien N. P. Martel, Gordon Wetzstein, Henry ArguelloICCV 2021 · 被引用 27 次
相关 Paper
- Lens Parameter Estimation for Realistic Depth of Field ModelingDominique Piché-Meunier, Yannick Hold-Geoffroy, Jianming Zhang, Jean-François LalondeICCV 2023 · 被引用 3 次
- MonoCD: Monocular 3D Object Detection with Complementary DepthsLongfei Yan, Pei Yan, Shengzhou Xiong, Xuanyu Xiang 等CVPR 2024 · 被引用 52 次
- Deep Depth From Aberration MapMasako Kashiwagi, Nao Mishima, Tatsuo Kozakaya, Shinsaku HiuraICCV 2019 · 被引用 9 次
- MonoDTR: Monocular 3D Object Detection with Depth-Aware TransformerKuan-Chih Huang, Tsung-Han Wu, Hung-Ting Su, Winston H. HsuCVPR 2022 · 被引用 199 次
- Deep Optics for Single-Shot High-Dynamic-Range ImagingChristopher A. Metzler, Hayato Ikoma, Yifan Peng, Gordon WetzsteinCVPR 2020
