Transferring to Real-World Layouts: A Depth-aware Framework for Scene Adaptation
Mu Chen, Zhedong Zheng, Yi Yang
Abstract
Scene segmentation via unsupervised domain adaptation (UDA) enables the transfer of knowledge acquired from source synthetic data to real-world target data, which largely reduces the need for manual pixel-level annotations in the target domain. To facilitate domain-invariant feature learning, existing methods typically mix data from both the source domain and target domain by simply copying and pasting pixels. Such vanilla methods are usually suboptimal since they do not take into account how well the mixed layouts correspond to real-world scenarios. Real-world scenarios are with an inherent layout. Real-world scenarios are with an inherent layout. We observe that semantic categories, such as sidewalks, buildings, and sky, display relatively consistent depth distributions, and could be clearly distinguished in a depth map. The model suffers from confusion in predicting the target domain due to the unrealistic mixing. For instance, it is not reasonable to directly paste the near "pedestrian" pixels into the remote "sky" area. Based on such observation, we propose a depth-aware framework to explicitly leverage depth estimation to mix categories and facilitate two complementary tasks, i.e., segmentation and depth learning in an end-to-end manner. In particular, the framework contains a Depth-guided Contextual Filter (DCF) for data augmentation and a cross-task encoder for contextual learning. DCF simulates the real-world layouts, while the cross-task encoder further adaptively fuses the complementing features between two tasks. Besides, several public datasets do not provide depth annotation. Therefore, we leverage the off-the-shelf depth estimation network to obtain the pseudo depth. Extensive experiments show that our methods, even with pseudo depth, achieve competitive performance, i.e., 77.7 mIoU on GTA→Cityscapes and 69.3 mIoU on Synthia→Cityscapes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4becd2c8-d45c-47c2-b6e5-8af67315bb99Cited by top-tier papers2
- Dual-stream Feature Augmentation for Domain GeneralizationShanshan Wang, ALuSi, Xun Yang, Ke Xu et al.ACM MM 2024 · 12 citations
- DiffVsgg: Diffusion-Driven Online Video Scene Graph GenerationMu Chen, Liulei Li, Wenguan Wang, Yi YangCVPR 2025
Builds on47
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Confidence Regularized Self-TrainingYang Zou, Zhiding Yu, Xiaofeng Liu, B. V. K. Vijaya Kumar et al.ICCV 2019 · 901 citations
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas et al.ICML 2020 · 651 citations
- DAFormer: Improving Network Architectures and Training Strategies for Domain-Adaptive Semantic SegmentationLukas Hoyer, Dengxin Dai, Luc Van GoolCVPR 2022 · 562 citations
Related papers
- DADA: Depth-Aware Domain Adaptation in Semantic SegmentationTuan-Hung Vu, Himalaya Jain, Maxime Bucher, Matthieu Cord et al.ICCV 2019 · 202 citations
- Geometry-Aware Network for Domain Adaptive Semantic SegmentationYinghong Liao, Wending Zhou, Xu Yan, Zhen Li et al.AAAI 2023 · 9 citations
- Domain Adaptive Semantic Segmentation with Self-Supervised Depth EstimationQin Wang, Dengxin Dai, Lukas Hoyer, Luc Van Gool et al.ICCV 2021 · 167 citations
- Unsupervised Domain Adaptation for Semantic Segmentation using Depth DistributionQuanliang Wu, Huajun LiuNeurIPS 2022 · 8 citations
- Exploring High-quality Target Domain Information for Unsupervised Domain Adaptive Semantic SegmentationJunjie Li, Zilei Wang, Yuan Gao, Xiaoming HuACM MM 2022 · 23 citations
