BridgeNet: A Joint Learning Network of Depth Map Super-Resolution and Monocular Depth Estimation
Qi Tang, Runmin Cong, Ronghui Sheng, Lingzhi He, Dan Zhang, Yao Zhao, Sam Kwong
Abstract
Depth map super-resolution is a task with high practical application requirements in the industry. Existing color-guided depth map super-resolution methods usually necessitate an extra branch to extract high-frequency detail information from RGB image to guide the low-resolution depth map reconstruction. However, because there are still some differences between the two modalities, direct information transmission in the feature dimension or edge map dimension cannot achieve satisfactory result, and may even trigger texture copying in areas where the structures of the RGB-D pair are inconsistent. Inspired by the multi-task learning, we propose a joint learning network of depth map super-resolution (DSR) and monocular depth estimation (MDE) without introducing additional supervision labels. For the interaction of two subnetworks, we adopt a differentiated guidance strategy and design two bridges correspondingly. One is the high-frequency attention bridge (HABdg) designed for the feature encoding process, which learns the high-frequency information of the MDE task to guide the DSR task. The other is the content guidance bridge (CGBdg) designed for the depth map reconstruction process, which provides the content guidance learned from DSR task for MDE task. The entire network architecture is highly portable and can provide a paradigm for associating the DSR and MDE tasks. Extensive experiments on benchmark datasets demonstrate that our method achieves competitive performance. Our code and models are available at https://rmcong.github.io/proj_BridgeNet.html.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ab46463c-bf0a-4e1f-aa8e-97510cb14b5fCited by top-tier papers11
- Discrete Cosine Transform Network for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Shuang Xu, Zudi Lin et al.CVPR 2022 · 120 citations
- SGNet: Structure Guided Network via Gradient-Frequency Awareness for Depth Map Super-resolutionZhengxue Wang, Zhiqiang Yan, Jian YangAAAI 2024 · 64 citations
- Spherical Space Feature Decomposition for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Xiang Gu, Chengli Tan et al.ICCV 2023 · 55 citations
- Recurrent Structure Attention Guidance for Depth Super-resolutionJiayi Yuan, Haobo Jiang, Xiang Li, Jianjun Qian et al.AAAI 2023 · 33 citations
- Symmetric Uncertainty-Aware Feature Transmission for Depth Super-ResolutionWuxuan Shi, Mang Ye, Bo DuACM MM 2022 · 23 citations
Builds on5
- Guided Image-to-Image Translation With Bi-Directional Feature TransformationBadour Albahar, Jia-Bin HuangICCV 2019 · 102 citations
- Guided Super-Resolution As Pixel-to-Pixel TransformationRiccardo de Lutio, Stefano D'Aronco, Jan Dirk Wegner, Konrad SchindlerICCV 2019 · 78 citations
- Joint Super-Resolution and Alignment of Tiny FacesYu Yin, Joseph P. Robinson, Yulun Zhang, Yun FuAAAI 2020 · 38 citations
- Towards Fast and Accurate Real-World Depth Super-Resolution: Benchmark Dataset and BaselineLingzhi He, Hongguang Zhu, Feng Li, Huihui Bai et al.CVPR 2021
- Learning Scene Structure Guidance via Cross-Task Knowledge Transfer for Single Depth Super-ResolutionBaoli Sun, Xinchen Ye, Baopu Li, Haojie Li et al.CVPR 2021
Related papers
- Structure Flow-Guided Network for Real Depth Super-resolutionJiayi Yuan, Haobo Jiang, Xiang Li, Jianjun Qian et al.AAAI 2023 · 17 citations
- Pyramid Dual Domain Injection Network for Pan-sharpeningXuanhua He, Keyu Yan, Rui Li, Chengjun Xie et al.ICCV 2023 · 15 citations
- Feedback Network for Mutually Boosted Stereo Image Super-Resolution and Disparity EstimationQinyan Dai, Juncheng Li, Qiaosi Yi, Faming Fang et al.ACM MM 2021 · 68 citations
- MonoMVSNet: Monocular Priors Guided Multi-View Stereo NetworkJianfei Jiang, Qiankun Liu, Haochen Yu, Hongyuan Liu et al.ICCV 2025 · 3 citations
- Unsupervised High-Resolution Depth Learning From Videos With Dual NetworksJunsheng Zhou, Yuwang Wang, Kaihuai Qin, Wenjun ZengICCV 2019 · 77 citations
