MobiDepth: real-time depth estimation using on-device dual cameras
Jinrui Zhang, Huan Yang, Ju Ren, Deyu Zhang, Bangwen He, Ting Cao, Yuanchun Li, Yaoxue Zhang, Yunxin Liu
摘要
Real-time depth estimation is critical for the increasingly popular augmented reality and virtual reality applications on mobile devices. Yet existing solutions are insufficient as they require expensive depth sensors or motion of the device, or have a high latency. We propose MobiDepth, a real-time depth estimation system using the widely-available on-device dual cameras. While binocular depth estimation is a mature technique, it is challenging to realize the technique on commodity mobile devices due to the different focal lengths and unsynchronized frame flows of the on-device dual cameras and the heavy stereo-matching algorithm.
To address the challenges, MobiDepth integrates three novel techniques: 1) iterative field-of-view cropping, which crops the field-of-views of the dual cameras to achieve the equivalent focal lengths for accurate epipolar rectification; 2) heterogeneous camera synchronization, which synchronizes the frame flows captured by the dual cameras to avoid the displacement of moving objects across the frames in the same pair; 3) mobile GPU-friendly stereo matching, which effectively reduces the latency of stereo matching on a mobile GPU. We implement MobiDepth on multiple commodity mobile devices and conduct comprehensive evaluations. Experimental results show that MobiDepth achieves real-time depth estimation of 22 frames per second with a significantly reduced depth-estimation error compared with the baselines. Using MobiDepth, we further build an example application of 3D pose estimation, which significantly outperforms the state-of-the-art 3D pose-estimation method, reducing the pose-estimation latency and error by up to 57.1% and 29.5%, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Asymmetric Dual-Lens Video DeblurringZeyu Xiao, Xinchao WangNeurIPS 2025 · 被引用 2 次
- Efficient Depth Estimation for Unstable Stereo Camera Systems on AR GlassesYongfan Liu, Hyoukjun KwonCVPR 2025
它引用的顶会 Paper5
- Heimdall: mobile GPU coordination platform for augmented reality applicationsJuheon Yi, Youngki LeeMobiCom 2020 · 被引用 71 次
- Romou: rapidly generate high-performance tensor kernels for mobile GPUsRendong Liang, Ting Cao, Jicheng Wen, Manni Wang 等MobiCom 2022 · 被引用 13 次
- HITNet: Hierarchical Iterative Tile Refinement Network for Real-time Stereo MatchingVladimir Tankovich, Christian Hane, Yinda Zhang, Adarsh Kowdle 等CVPR 2021
- Lite-HRNet: A Lightweight High-Resolution NetworkChangqian Yu, Bin Xiao, Changxin Gao, Lu Yuan 等CVPR 2021
- HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose EstimationBowen Cheng, Bin Xiao, Jingdong Wang, Honghui Shi 等CVPR 2020
相关 Paper
- FlashDepth: Real-Time Streaming Video Depth Estimation at 2K ResolutionGene Chou, Wenqi Xian, Guandao Yang, Mohamed Abdelfattah 等ICCV 2025 · 被引用 1 次
- EasyREG: Easy Depth-Based Markerless Registration and Tracking using Augmented Reality Device for Surgical GuidanceYue Yang, Christoph Leuze, Brian A. Hargreaves, Bruce Daniel 等IEEE VR 2026 · 被引用 4 次
- Lite Pose: Efficient Architecture Design for 2D Human Pose EstimationYihan Wang, Muyang Li, Han Cai, Wei-Ming Chen 等CVPR 2022 · 被引用 117 次
- Learning Single Camera Depth Estimation Using Dual-PixelsRahul Garg, Neal Wadhwa, Sameer Ansari, Jonathan T. BarronICCV 2019 · 被引用 123 次
- RePoseD: Efficient Relative Pose Estimation With Known Depth InformationYaqing Ding, Viktor Kocur, Václav Vávra, Zuzana Berger Haladová 等ICCV 2025 · 被引用 2 次
