Domain-Agnostic Mutual Prompting for Unsupervised Domain Adaptation
Zhekai Du, Xinyao Li, Fengling Li, Ke Lu, Lei Zhu, Jingjing Li
Abstract
Conventional Unsupervised Domain Adaptation (UDA) strives to minimize distribution discrepancy between domains, which neglects to harness rich semantics from data and struggles to handle complex domain shifts. A promising technique is to leverage the knowledge of large-scale pretrained vision-language models for more guided adaptation. Despite some endeavors, current methods often learn textual prompts to embed domain semantics for source and target domains separately and perform classification within each domain, limiting cross-domain knowledge transfer. Moreover, prompting only the language branch lacks flexibility to adapt both modalities dynamically. To bridge this gap, we propose Domain-Agnostic Mutual Prompting (DAMP) to exploit domain-invariant semantics by mutually aligning visual and textual embeddings. Specifically, the image contextual information is utilized to prompt the language branch in a domain-agnostic and instanceconditioned way. Meanwhile, visual prompts are imposed based on the domain-agnostic textual prompt to elicit domain-invariant visual embeddings. These two branches of prompts are learned mutually with a cross-attention module and regularized with a semantic-consistency loss and an instance-discrimination contrastive loss. Experiments on three UDA benchmarks demonstrate the superiority of DAMP over state-of-the-art approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b31da7b3-909c-4836-9e50-111c212c678eCited by top-tier papers17
- Enhancing Domain Adaptation through Prompt Gradient AlignmentViet Hoang Phan, Tung Lam Tran, Quyen Tran, Trung LeNeurIPS 2024 · 18 citations
- Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain AdaptationKuanghong Liu, Jin Wang, Kangjian He, Dan Xu et al.AAAI 2025 · 4 citations
- STraj: Self-training for Bridging the Cross-Geography Gap in Trajectory PredictionZhanwei Zhang, Minghao Chen, Zhihong Gu, Xinkui Zhao et al.AAAI 2025 · 2 citations
- Hybrid-Tta: Continual Test-Time Adaptation Via Dynamic Domain Shift DetectionHyewon Park, Hyejin Park, Jueun Ko, Dongbo MinICCV 2025 · 2 citations
- CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain AdaptationMainak Singha, Sarthak Mehrotra, Paolo Casari, Subhasis Chaudhuri et al.CVPR 2026 · 2 citations
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen et al.ICML 2021 · 5,401 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
Related papers
- Prompt-Based Distribution Alignment for Unsupervised Domain AdaptationShuanghao Bai, Min Zhang, Wanqi Zhou, Siteng Huang et al.AAAI 2024 · 103 citations
- Domain-aware Visual Context Prompt for Multi-Source Domain AdaptationYuwu Lu, Haoyu Huang, Xue HuACM MM 2025 · 1 citation
- CLIP2UDA: Making Frozen CLIP Reward Unsupervised Domain Adaptation in 3D Semantic SegmentationYao Wu, Mingwei Xing, Yachao Zhang, Yuan Xie et al.ACM MM 2024 · 12 citations
- End-to-End Knowledge Distillation for Unsupervised Domain Adaptation with Large Vision-language ModelsYangtao Wang, Xingwei Deng, Yanzhao Xie, Weilong Peng et al.AAAI 2026 · 1 citation
- CLIP2Pose: Frozen CLIP as Semantic Guide for Domain Adaptive Pose EstimationJiawen Li, Fei Jiang, Dandan Zhu, Jinxin Shi et al.AAAI 2026
