Learning Scene Structure Guidance via Cross-Task Knowledge Transfer for Single Depth Super-Resolution
Baoli Sun, Xinchen Ye, Baopu Li, Haojie Li, Zhihui Wang, Rui Xu
Abstract
Existing color-guided depth super-resolution (DSR) approaches require paired RGB-D data as training samples where the RGB image is used as structural guidance to recover the degraded depth map due to their geometrical similarity. However, the paired data may be limited or expensive to be collected in actual testing environment. Therefore, we explore for the first time to learn the cross-modality knowledge at training stage, where both RGB and depth modalities are available, but test on the target dataset, where only single depth modality exists. Our key idea is to distill the knowledge of scene structural guidance from RG-B modality to the single DSR task without changing its network architecture. Specifically, we construct an auxiliary depth estimation (DE) task that takes an RGB image as input to estimate a depth map, and train both D-SR task and DE task collaboratively to boost the performance of DSR. Upon this, a cross-task interaction module is proposed to realize bilateral cross-task knowledge transfer. First, we design a cross-task distillation scheme that encourages DSR and DE networks to learn from each other in a teacher-student role-exchanging fashion. Then, we advance a structure prediction (SP) task that provides extra structure regularization to help both DSR and DE networks learn more informative structure representations for depth recovery. Extensive experiments demonstrate that our scheme achieves superior performance in comparison with other DSR methods. Our code available at: https: //github.com/Sunbaoli/dsr-distillation .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a1a656fe-2ea5-4696-a92f-767641b32ceaCited by top-tier papers14
- Discrete Cosine Transform Network for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Shuang Xu, Zudi Lin et al.CVPR 2022 · 120 citations
- SGNet: Structure Guided Network via Gradient-Frequency Awareness for Depth Map Super-resolutionZhengxue Wang, Zhiqiang Yan, Jian YangAAAI 2024 · 64 citations
- Spherical Space Feature Decomposition for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Xiang Gu, Chengli Tan et al.ICCV 2023 · 55 citations
- BridgeNet: A Joint Learning Network of Depth Map Super-Resolution and Monocular Depth EstimationQi Tang, Runmin Cong, Ronghui Sheng, Lingzhi He et al.ACM MM 2021 · 47 citations
- Learning Graph Regularisation for Guided Super-ResolutionRiccardo de Lutio, Alexander Becker, Stefano D'Aronco, Stefania Russo et al.CVPR 2022 · 40 citations
Builds on5
- Guided Image-to-Image Translation With Bi-Directional Feature TransformationBadour Albahar, Jia-Bin HuangICCV 2019 · 102 citations
- Guided Super-Resolution As Pixel-to-Pixel TransformationRiccardo de Lutio, Stefano D'Aronco, Jan Dirk Wegner, Konrad SchindlerICCV 2019 · 78 citations
- UM-Adapt: Unsupervised Multi-Task Adaptation Using Adversarial Cross-Task DistillationJogendra Nath Kundu, Nishank Lakkakula, Venkatesh Babu RadhakrishnanICCV 2019 · 62 citations
- SDC-Depth: Semantic Divide-and-Conquer Network for Monocular Depth EstimationLijun Wang, Jianming Zhang, Oliver Wang, Zhe Lin et al.CVPR 2020
- Structure-Preserving Super Resolution With Gradient GuidanceCheng Ma, Yongming Rao, Yean Cheng, Ce Chen et al.CVPR 2020
Related papers
- Knowledge As Priors: Cross-Modal Knowledge Generalization for Datasets Without Superior KnowledgeLong Zhao, Xi Peng, Yuxiao Chen, Mubbasir Kapadia et al.CVPR 2020
- Symmetric Uncertainty-Aware Feature Transmission for Depth Super-ResolutionWuxuan Shi, Mang Ye, Bo DuACM MM 2022 · 23 citations
- Structure Flow-Guided Network for Real Depth Super-resolutionJiayi Yuan, Haobo Jiang, Xiang Li, Jianjun Qian et al.AAAI 2023 · 17 citations
- RGB-Multispectral Matching: Dataset, Learning Methodology, EvaluationFabio Tosi, Pierluigi Zama Ramirez, Matteo Poggi, Samuele Salti et al.CVPR 2022 · 5 citations
- Cross-Domain 3D Hand Pose Estimation with Dual ModalitiesQiuxia Lin, Linlin Yang, Angela YaoCVPR 2023
