Cross-Class Domain Adaptive Semantic Segmentation with Visual Language Models
Wenqi Ren, Ruihao Xia, Meng Zheng, Ziyan Wu, Yang Tang, Nicu Sebe
摘要
This paper addresses the issue of cross-class domain adaptation (CCDA) in semantic segmentation, where the target domain contains both shared and novel classes that are either unlabeled or unseen in the source domain. This problem is challenging, as the absence of labels for novel classes hampers the effective solutions of both cross-domain and cross-class problems. Since Visual Language Models (VLMs) have exhibited impressive generalization across diverse data distributions and are capable of generating zero-shot predictions without requiring task-specific training examples, we propose a label alignment method by leveraging VLMs to relabel pseudo labels for novel classes. Considering that VLMs typically provide only image-level predictions, we embed a two-stage method to enable fine-grained semantic segmentation and design a threshold based on the uncertainty of pseudo labels to exclude noisy VLM predictions. To further augment the supervision of novel classes, we devise memory banks with an adaptive update scheme to effectively manage accurate VLM predictions, which are then resampled to increase the sampling probability of novel classes. Through comprehensive experiments, we demonstrate the effectiveness and versatility of our proposed method across various CCDA scenarios.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- SemiDAViL: Semi-supervised Domain Adaptation with Vision-Language Guidance for Semantic SegmentationHritam Basak, Zhaozheng YinCVPR 2025
- Domain Adaptive Hashing Retrieval via VLM Assisted Pseudo-Labeling and Dual Space AdaptationJingyao Li, Zhanshan Li, Shuai LüNeurIPS 2025 · 被引用 1 次
- DA-Ada: Learning Domain-Aware Adapter for Domain Adaptive Object DetectionHaochen Li, Rui Zhang, Hantao Yao, Xin Zhang 等NeurIPS 2024 · 被引用 20 次
- Prompt-Based Distribution Alignment for Unsupervised Domain AdaptationShuanghao Bai, Min Zhang, Wanqi Zhou, Siteng Huang 等AAAI 2024 · 被引用 103 次
- Vision-Language Model Guided Source-Free Domain Adaptation via Optimal TransportShuo Han, Xu Tang, Jingjing Ma, Xiangrong ZhangCVPR 2026
