Self-Supervised Deep Depth Denoising
Vladimiros Sterzentsenko, Leonidas Saroglou, Anargyros Chatzitofis, Spiros Thermos, Nikolaos Zioulis, Alexandros Doumanoglou, Dimitrios Zarpalas, Petros Daras
Abstract
Depth perception is considered an invaluable source of information for various vision tasks. However, depth maps acquired using consumer-level sensors still suffer from non-negligible noise. This fact has recently motivated researchers to exploit traditional filters, as well as the deep learning paradigm, in order to suppress the aforementioned non-uniform noise, while preserving geometric details. Despite the effort, deep depth denoising is still an open challenge mainly due to the lack of clean data that could be used as ground truth. In this paper, we propose a fully convolutional deep autoencoder that learns to denoise depth maps, surpassing the lack of ground truth data. Specifically, the proposed autoencoder exploits multiple views of the same scene from different points of view in order to learn to suppress noise in a self-supervised end-to-end manner using depth and color information during training, yet only depth during inference. To enforce self-supervision, we leverage a differentiable rendering technique to exploit photometric supervision, which is further regularized using geometric and surface priors. As the proposed approach relies on raw data acquisition, a large RGB-D corpus is collected using Intel RealSense sensors. Complementary to a quantitative evaluation, we demonstrate the effectiveness of the proposed self-supervised denoising approach on established 3D reconstruction applications. Code is avalable at https://github.com/VCL3D/DeepDepthDenoising
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor ControlZichen Jeff Cui, Hengkai Pan, Aadhithya Iyer, Siddhant Haldar et al.NeurIPS 2024 · 61 citations
- Geometry-aware Two-scale PIFu Representation for Human ReconstructionZheng Dong, Ke Xu, Ziheng Duan, Hujun Bao et al.NeurIPS 2022 · 17 citations
- Token Boosting for Robust Self-Supervised Visual Transformer Pre-trainingTianjiao Li, Lin Geng Foo, Ping Hu, Xindi Shang et al.CVPR 2023
- Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-ResolutionHuan Zheng, Wencheng Han, Jianbing ShenCVPR 2025
Related papers
- UnsupervisedR&R: Unsupervised Point Cloud Registration via Differentiable RenderingMohamed El Banani, Luya Gao, Justin JohnsonCVPR 2021
- Ponder: Point Cloud Pre-training via Neural RenderingDi Huang, Sida Peng, Tong He, Honghui Yang et al.ICCV 2023 · 55 citations
- Multi-view Self-supervised Disentanglement for General Image DenoisingHao Chen, Chenyuan Qu, Yu Zhang, Chen Chen et al.ICCV 2023 · 15 citations
- Single Image Depth Prediction With Wavelet DecompositionMichaël Ramamonjisoa, Michael Firman, Jamie Watson, Vincent Lepetit et al.CVPR 2021
- Fully Self-Supervised Depth Estimation from Defocus ClueHaozhe Si, Bin Zhao, Dong Wang, Yunpeng Gao et al.CVPR 2023
