Self-supervised Monocular Underwater Depth Recovery, Image Restoration, and a Real-sea Video Dataset
Nisha Varghese, Ashish Kumar, A. N. Rajagopalan
Abstract
Underwater (UW) depth estimation and image restoration is a challenging task due to its fundamental ill-posedness and the unavailability of real large-scale UW-paired datasets. UW depth estimation has been attempted before by utilizing either the haze information present or the geometry cue from stereo images or the adjacent frames in a video. To obtain improved estimates of depth from a single UW image, we propose a deep learning (DL) method that utilizes both haze and geometry during training. By harnessing the physical model for UW image formation in conjunction with the view-synthesis constraint on neighboring frames in monocular videos, we perform disentanglement of the input image to also get an estimate of the scene radiance. The proposed method is completely self-supervised and simultaneously outputs the depth map and the restored image in real-time (55 fps). We call this first-ever Underwater Self-supervised deep learning network for simultaneous Recovery of Depth and Image as USe-ReDI-Net. To facilitate monocular self-supervision, we collected a Dataset of Real-world Underwater Videos of Artifacts (DRUVA) in shallow sea waters. DRUVA is the first UW video dataset that contains video sequences of 20 different submerged artifacts with almost full azimuthal coverage of each artifact. Extensive experiments on our DRUVA dataset and other UW datasets establish the superiority of our proposed USe-ReDI-Net over prior art for both UW depth and image recovery. The dataset DRUVA is available at https://github.com/nishavarghese15/DRUVA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Atlantis: Enabling Underwater Depth Estimation with Stable DiffusionFan Zhang, Shaodi You, Yu Li, Ying FuCVPR 2024 · 22 citations
- UVEB: A Large-scale Benchmark and Baseline Towards Real-World Underwater Video EnhancementYaofeng Xie, Lingwei Kong, Kai Chen, Ziqiang Zheng et al.CVPR 2024 · 20 citations
- NeuroPump: Simultaneous Geometric and Color Rectification for Underwater ImagesYue Guo, Haoxiang Liao, Haibin Ling, Bingyao HuangACM MM 2025 · 1 citation
- Beyond Appearance: Camouflaged Object Detection via Geometric StructureJinyu Han, Changguang Wu, Fuming Sun, Jinhui TangCVPR 2026
- UVLM: Benchmarking Video Language Model for Underwater World UnderstandingXizhe Xue, Yang Zhou, Dawei Yan, Lijie Tao et al.AAAI 2026
Builds on2
Related papers
- Sea-ing in Low-lightNisha Varghese, A. N. RajagopalanCVPR 2025
- Learning to Remove Refractive Distortions from Underwater ImagesSimron Thapa, Nianyi Li, Jinwei YeICCV 2021 · 16 citations
- Dynamic Fluid Surface Reconstruction Using Deep Neural NetworkSimron Thapa, Nianyi Li, Jinwei YeCVPR 2020
- DeLightMono: Enhancing Self-Supervised Monocular Depth Estimation in Endoscopy by Decoupling Uneven IlluminationMingyang Ou, Haojin Li, Yifeng Zhang, Ke Niu et al.AAAI 2026
- DäRF: Boosting Radiance Fields from Sparse Input Views with Monocular Depth AdaptationJiuhn Song, Seonghoon Park, Honggyu An, Seokju Cho et al.NeurIPS 2023 · 11 citations
