Single Image Depth Prediction With Wavelet Decomposition
Michaël Ramamonjisoa, Michael Firman, Jamie Watson, Vincent Lepetit, Daniyar Turmukhambetov
Abstract
We present a novel method for predicting accurate depths from monocular images with high efficiency. This optimal efficiency is achieved by exploiting wavelet decomposition, which is integrated in a fully differentiable encoder-decoder architecture. We demonstrate that we can reconstruct high-fidelity depth maps by predicting sparse wavelet coefficients. In contrast with previous works, we show that wavelet coefficients can be learned without direct supervision on coefficients. Instead we supervise only the final depth image that is reconstructed through the inverse wavelet transform. We additionally show that wavelet coefficients can be learned in fully self-supervised scenarios, without access to ground-truth depth. Finally, we apply our method to different state-of-the-art monocular depth estimation models, in each case giving similar or better results compared to the original model, while requiring less than half the multiplyadds in the decoder network.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ea08602f-433e-4e36-861a-561716df2f18Cited by top-tier papers6
- DGECN: A Depth-Guided Edge Convolutional Network for End-to-End 6D Pose EstimationTuo Cao, Fei Luo, Yanping Fu, Wenxiao Zhang et al.CVPR 2022 · 43 citations
- Deep Depth from Focus with Differential Focus VolumeFengting Yang, Xiaolei Huang, Zihan ZhouCVPR 2022 · 31 citations
- AccuMO: Accuracy-Centric Multitask Offloading in Edge-Assisted Mobile Augmented RealityZ. Jonny Kong, Qiang Xu, Jiayi Meng, Y. Charlie HuMobiCom 2023 · 21 citations
- Adversarial Training of Self-supervised Monocular Depth Estimation against Physical-World AttacksZhiyuan Cheng, James Liang, Guanhong Tao, Dongfang Liu et al.ICLR 2023 · 6 citations
- FAAR: Efficient Frequency-Aware Multi-Task Fine-Tuning via Automatic Rank SelectionMaxime Fontana, Michael W. Spratling, Miaojing ShiCVPR 2026
Builds on8
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Depth From Videos in the Wild: Unsupervised Monocular Depth Learning From Unknown CamerasAriel Gordon, Hanhan Li, Rico Jonschkowski, Anelia AngelovaICCV 2019 · 397 citations
- Self-Supervised Monocular Depth HintsJamie Watson, Michael Firman, Gabriel J. Brostow, Daniyar TurmukhambetovICCV 2019 · 287 citations
- Self-Supervised Learning With Geometric Constraints in Monocular Video: Connecting Flow, Depth, and CameraYuhua Chen, Cordelia Schmid, Cristian SminchisescuICCV 2019 · 265 citations
- Wavelet Domain Style Transfer for an Effective Perception-Distortion Tradeoff in Single Image Super-ResolutionXin Deng, Ren Yang, Mai Xu, Pier Luigi DragottiICCV 2019 · 87 citations
Related papers
- R-MSFM: Recurrent Multi-Scale Feature Modulation for Monocular Depth EstimatingZhongkai Zhou, Xinnan Fan, Pengfei Shi, Yuanxue XinICCV 2021 · 150 citations
- Fully Self-Supervised Depth Estimation from Defocus ClueHaozhe Si, Bin Zhao, Dong Wang, Yunpeng Gao et al.CVPR 2023
- Unsupervised High-Resolution Depth Learning From Videos With Dual NetworksJunsheng Zhou, Yuwang Wang, Kaihuai Qin, Wenjun ZengICCV 2019 · 77 citations
- Self-Supervised Deep Depth DenoisingVladimiros Sterzentsenko, Leonidas Saroglou, Anargyros Chatzitofis, Spiros Thermos et al.ICCV 2019 · 47 citations
- SQLdepth: Generalizable Self-Supervised Fine-Structured Monocular Depth EstimationYouhong Wang, Yunji Liang, Hao Xu, Shaohui Jiao et al.AAAI 2024 · 60 citations
