CDAC: Cross-domain Attention Consistency in Transformer for Domain Adaptive Semantic Segmentation
Kaihong Wang, Donghyun Kim, Rogério Feris, Margrit Betke
Abstract
While transformers have greatly boosted performance in semantic segmentation, domain adaptive transformers are not yet well explored. We identify that the domain gap can cause discrepancies in self-attention. Due to this gap, the transformer attends to spurious regions or pixels, which deteriorates accuracy on the target domain. We propose Cross-Domain Attention Consistency (CDAC), to perform adaptation on attention maps using cross-domain attention layers that share features between source and target domains. Specifically, we impose consistency between predictions from cross-domain attention and self-attention modules to encourage similar distributions across domains in both the attention and output of the model, i.e., attention-level and output-level alignment. We also enforce consistency in attention maps between different augmented views to further strengthen the attention-based alignment. Combining these two components, CDAC mitigates the discrepancy in attention maps across domains and further boosts the performance of the transformer under unsupervised domain adaptation settings. Our method is evaluated on various widely used benchmarks and outperforms the state-of-the-art baselines, including GTAV-to-Cityscapes by 1.3 and 1.5 percent point (pp) and Synthia-to-Cityscapes by 0.6 pp and 2.9 pp when combining with two competitive Transformer-based backbones, respectively. Our code will be publicly available at https://github.com/wangkaihong/CDAC.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ce4df40e-26c7-483a-948c-eaa9d46f6b8dCited by top-tier papers7
- Generalized Category Discovery under Domain Shift: A Frequency Domain PerspectiveWei Feng, Zongyuan GeNeurIPS 2025 · 9 citations
- EAGLE: Efficient Adaptive Geometry-based Learning in Cross-view UnderstandingThanh-Dat Truong, Utsav Prabhu, Dongyi Wang, Bhiksha Raj et al.NeurIPS 2024 · 7 citations
- Towards Robust Pseudo-Label Learning in Semantic Segmentation: An Encoding PerspectiveWangkai Li, Rui Sun, Zhaoyang Li, Tianzhu ZhangNeurIPS 2025 · 5 citations
- In-Context Policy Adaptation via Cross-Domain Skill DiffusionMinjong Yoo, Woo Kyung Kim, Honguk WooAAAI 2025 · 3 citations
- ERF: A Benchmark Dataset for Robust Semantic Segmentation Under Extreme Rainfall ConditionsXin Yang, Xin Zhang, Xinchao WangAAAI 2025 · 3 citations
Builds on23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty et al.NeurIPS 2021 · 2,985 citations
Related papers
- PixMatch: Unsupervised Domain Adaptation via Pixelwise Consistency TrainingLuke Melas-Kyriazi, Arjun K. ManraiCVPR 2021
- Addressing Domain Gap via Content Invariant Representation for Semantic SegmentationLi Gao, Lefei Zhang, Qian ZhangAAAI 2021 · 23 citations
- Pixel-Level Cycle Association: A New Perspective for Domain Adaptive Semantic SegmentationGuoliang Kang, Yunchao Wei, Yi Yang, Yueting Zhuang et al.NeurIPS 2020 · 124 citations
- Differential Treatment for Stuff and Things: A Simple Unsupervised Domain Adaptation Method for Semantic SegmentationZhonghao Wang, Mo Yu, Yunchao Wei, Rogério Feris et al.CVPR 2020
- PiPa: Pixel- and Patch-wise Self-supervised Learning for Domain Adaptative Semantic SegmentationMu Chen, Zhedong Zheng, Yi Yang, Tat-Seng ChuaACM MM 2023 · 65 citations
