What Synthesis Is Missing: Depth Adaptation Integrated With Weak Supervision for Indoor Scene Parsing
Keng-Chi Liu, Yi-Ting Shen, Jan Klopp, Liang-Gee Chen
摘要
Scene Parsing is a crucial step to enable autonomous systems to understand and interact with their surroundings. Supervised deep learning methods have made great progress in solving scene parsing problems, however, come at the cost of laborious manual pixel-level annotation. To alleviate this effort synthetic data as well as weak supervision have both been investigated. Nonetheless, synthetically generated data still suffers from severe domain shift while weak labels are often imprecise. Moreover, most existing works for weakly supervised scene parsing are limited to salient foreground objects. The aim of this work is hence twofold: Exploit synthetic data where feasible and integrate weak supervision where necessary. More concretely, we address this goal by utilizing depth as transfer domain because its synthetic-to-real discrepancy is much lower than for color. At the same time, we perform weak localization from easily obtainable image level labels and integrate both using a novel contour-based scheme. Our approach is implemented as a teacher-student learning framework to solve the transfer learning problem by generating a pseudo ground truth. Using only depth-based adaptation, this approach already outperforms previous transfer learning approaches on the popular indoor scene parsing SUN RGB-D dataset. Our proposed two-stage integration more than halves the gap towards fully supervised methods when compared to previous state-of-the-art in transfer learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Toward Real-World High-Precision Image Matting and SegmentationHaipeng Zhou, Zhaohu Xing, Hongqiu Wang, Jun Ma 等AAAI 2026
- Transferring to Real-World Layouts: A Depth-aware Framework for Scene AdaptationMu Chen, Zhedong Zheng, Yi YangACM MM 2024 · 被引用 19 次
- DADA: Depth-Aware Domain Adaptation in Semantic SegmentationTuan-Hung Vu, Himalaya Jain, Maxime Bucher, Matthieu Cord 等ICCV 2019 · 被引用 202 次
- Semi-Supervised Stereo-Based 3D Object Detection via Cross-View ConsensusWenhao Wu, Hau-San Wong, Si WuCVPR 2023
- 3D-to-2D Distillation for Indoor Scene ParsingZhengzhe Liu, Xiaojuan Qi, Chi-Wing FuCVPR 2021
