Real2Sim2Real: RetinalDepth-64K for Depth Estimation in Posterior Segment Ophthalmic Surgery
Bingwen Dong, Gan Liu, Xiaoxi Lu, Guangcheng Chen, Jialu Zhang, Yan Hu, Xiaoqing Zhang, Jiang Liu
摘要
Accurate depth estimation is crucial for 3D reconstruction and precise navigation in posterior segment ophthalmic surgery. However, acquiring annotated data remains challenging due to the impracticality of depth sensors under surgical microscopes. To overcome this limitation, we introduce RetinalDepth, a novel synthetic dataset comprising 44,800 temporal stereo image pairs across 896 diverse scenes for posterior segment surgery, developed via a Real2Sim2Real pipeline. In the Real-to-Sim phase, we model anatomically accurate eyes and instruments in Blender, integrating ultra-wide-field retinal textures, refractive aqueous humor, and dynamic trajectories, enhanced by post-processing for realism. RetinalDepth provides synchronized left and right RGB frames with pixel-perfect depth maps, surface normals, instrument segmentation masks, and camera parameters, enabling robust training of monocular, stereo, and video depth estimation models. In the Sim-to-Real phase, we introduce temporal depth variance, a novel metric for quantifying the stability of frameto-frame depth estimation. Fine-tuning on RetinalDepth significantly boosts model performance on both synthetic and real surgical videos, enhancing generalization, surgical precision, and novice training. As the first synthetic benchmark for posterior segment surgery, RetinalDepth bridges the sim-to-real gap in ophthalmic computer vision. Project page: https://retinaldepth.github.io/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao 等NeurIPS 2024 · 被引用 2,305 次
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu 等CVPR 2024 · 被引用 847 次
- Metric3D: Towards Zero-shot Metric 3D Prediction from A Single ImageWei Yin, Chi Zhang, Hao Chen, Zhipeng Cai 等ICCV 2023 · 被引用 388 次
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii 等CVPR 2024 · 被引用 302 次
- Repurposing Diffusion-Based Image Generators for Monocular Depth EstimationBingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger 等CVPR 2024
相关 Paper
- Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic SurgeryMing Hu, Zhengdi Yu, Feilong Tang, Kaiwen Chen 等NeurIPS 2025 · 被引用 1 次
- Seeing and Seeing Through the Glass: Real and Synthetic Data for Multi-Layer Depth EstimationHongyu Wen, Yiming Zuo, Venkat Subramanian, Patrick Chen 等ICCV 2025 · 被引用 1 次
- OphCLIP: Hierarchical Retrieval-Augmented Learning for Ophthalmic Surgical Video-Language PretrainingMing Hu, Kun Yuan, Yaling Shen, Feilong Tang 等ICCV 2025 · 被引用 1 次
- EgoGen: An Egocentric Synthetic Data GeneratorGen Li, Kaifeng Zhao, Siwei Zhang, Xiaozhong Lyu 等CVPR 2024
- Toward Practical Monocular Indoor Depth EstimationCho-Ying Wu, Jialiang Wang, Michael Hall, Ulrich Neumann 等CVPR 2022 · 被引用 68 次
