What Synthesis Is Missing: Depth Adaptation Integrated With Weak Supervision for Indoor Scene Parsing
Keng-Chi Liu, Yi-Ting Shen, Jan Klopp, Liang-Gee Chen
Abstract
Scene Parsing is a crucial step to enable autonomous systems to understand and interact with their surroundings. Supervised deep learning methods have made great progress in solving scene parsing problems, however, come at the cost of laborious manual pixel-level annotation. To alleviate this effort synthetic data as well as weak supervision have both been investigated. Nonetheless, synthetically generated data still suffers from severe domain shift while weak labels are often imprecise. Moreover, most existing works for weakly supervised scene parsing are limited to salient foreground objects. The aim of this work is hence twofold: Exploit synthetic data where feasible and integrate weak supervision where necessary. More concretely, we address this goal by utilizing depth as transfer domain because its synthetic-to-real discrepancy is much lower than for color. At the same time, we perform weak localization from easily obtainable image level labels and integrate both using a novel contour-based scheme. Our approach is implemented as a teacher-student learning framework to solve the transfer learning problem by generating a pseudo ground truth. Using only depth-based adaptation, this approach already outperforms previous transfer learning approaches on the popular indoor scene parsing SUN RGB-D dataset. Our proposed two-stage integration more than halves the gap towards fully supervised methods when compared to previous state-of-the-art in transfer learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- Toward Real-World High-Precision Image Matting and SegmentationHaipeng Zhou, Zhaohu Xing, Hongqiu Wang, Jun Ma et al.AAAI 2026
- Transferring to Real-World Layouts: A Depth-aware Framework for Scene AdaptationMu Chen, Zhedong Zheng, Yi YangACM MM 2024 · 19 citations
- DADA: Depth-Aware Domain Adaptation in Semantic SegmentationTuan-Hung Vu, Himalaya Jain, Maxime Bucher, Matthieu Cord et al.ICCV 2019 · 202 citations
- Semi-Supervised Stereo-Based 3D Object Detection via Cross-View ConsensusWenhao Wu, Hau-San Wong, Si WuCVPR 2023
- 3D-to-2D Distillation for Indoor Scene ParsingZhengzhe Liu, Xiaojuan Qi, Chi-Wing FuCVPR 2021
