DynAlign: Unsupervised Dynamic Taxonomy Alignment for Cross-Domain Segmentation
Han Sun, Rui Gong, Ismail Nejjar, Olga Fink
Abstract
Current unsupervised domain adaptation (UDA) methods for semantic segmentation typically assume identical class labels between the source and target domains. This assumption ignores the label-level domain gap, which is common in real-world scenarios, thus limiting their ability to identify finer-grained or novel categories without requiring extensive manual annotation. A promising direction to address this limitation lies in recent advancements in foundation models, which exhibit strong generalization abilities due to their rich prior knowledge. However, these models often struggle with domain-specific nuances and underrepresented fine-grained categories. To address these challenges, we introduce DynAlign , a framework that integrates UDA with foundation models to bridge both the image-level and label-level domain gaps. Our approach leverages prior semantic knowledge to align source categories with target categories that can be novel, more fine-grained, or named differently (e.g., 'vehicle' to 'car', 'truck', 'bus'). Foundation models are then employed for precise segmentation and category reassignment. To further enhance accuracy, we propose a knowledge fusion approach that dynamically adapts to varying scene contexts. DynAlign generates accurate predictions in a new target label space without requiring any manual annotations, allowing seamless adaptation to new taxonomies through either model retraining or direct inference. Experiments on the street scene semantic segmentation benchmarks GTA→Mapillary Vistas and GTA→IDD validate the effectiveness of our approach, achieving a significant improvement over existing methods. Our code is publically available at https://github.com/hansunhayden/DynAlign .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Language-driven Semantic SegmentationBoyi Li, Kilian Q. Weinberger, Serge J. Belongie, Vladlen Koltun et al.ICLR 2022 · 885 citations
- DAFormer: Improving Network Architectures and Training Strategies for Domain-Adaptive Semantic SegmentationLukas Hoyer, Dengxin Dai, Luc Van GoolCVPR 2022 · 562 citations
Related papers
- Differential Treatment for Stuff and Things: A Simple Unsupervised Domain Adaptation Method for Semantic SegmentationZhonghao Wang, Mo Yu, Yunchao Wei, Rogério Feris et al.CVPR 2020
- Universal Domain Adaptation for Semantic SegmentationSeun-An Choe, Keon-Hee Park, Jinwoo Choi, Gyeong-Moon ParkCVPR 2025
- DSP: Dual Soft-Paste for Unsupervised Domain Adaptive Semantic SegmentationLi Gao, Jing Zhang, Lefei Zhang, Dacheng TaoACM MM 2021 · 78 citations
- Coarse-To-Fine Domain Adaptive Semantic Segmentation With Photometric Alignment and Category-Center RegularizationHaoyu Ma, Xiangru Lin, Zifeng Wu, Yizhou YuCVPR 2021
- Category Dictionary Guided Unsupervised Domain Adaptation for Object DetectionShuai Li, Jianqiang Huang, Xian-Sheng Hua, Lei ZhangAAAI 2021 · 47 citations
