Domain-Conditioned Transformer for Fully Test-time Adaptation
Yushun Tang, Shuoshuo Chen, Jiyuan Jia, Yi Zhang, Zhihai He
摘要
Fully test-time adaptation aims to adapt a network model online based on sequential analysis of input samples during the inference stage. We observe that, when applying a transformer network model into a new domain, the self-attention profiles of image samples in the target domain deviate significantly from those in the source domain, which results in large performance degradation during domain changes. To address this important issue, we propose a new structure for the self-attention modules in the transformer. Specifically, we incorporate three domain-conditioning vectors, called domain conditioners, into the query, key, and value components of the self-attention module. We learn a network to generate these three domain conditioners from the class token at each transformer network layer. We find that, during fully online test-time adaptation, these domain conditioners at each transform network layer are able to gradually remove the impact of domain shift and largely recover the original self-attention profile. Our extensive experimental results demonstrate that the proposed domain-conditioned transformer significantly improves the online fully test-time domain adaptation performance and outperforms existing state-of-the-art methods by large margins. The code is available at https://github.com/yushuntang/DCT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 被引用 1,861 次
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen 等ICLR 2021 · 被引用 1,731 次
相关 Paper
- What, How, and When Should Object Detectors Update in Continually Changing Test Domains?Jayeon Yoo, Dongkwan Lee, Inseop Chung, Donghyun Kim 等CVPR 2024 · 被引用 10 次
- Neuro-Modulated Hebbian Learning for Fully Test-Time AdaptationYushun Tang, Ce Zhang, Heng Xu, Shuoshuo Chen 等CVPR 2023
- Back to Source: Open-Set Continual Test-Time Adaptation via Domain CompensationYingkai Yang, Chaoqi Chen, Hui HuangCVPR 2026 · 被引用 1 次
- Test-Time Adaptation via Self-Training with Nearest Neighbor InformationMinguk Jang, Sae-Young Chung, Hye Won ChungICLR 2023 · 被引用 12 次
- Decorate the Newcomers: Visual Domain Prompt for Continual Test Time AdaptationYulu Gan, Yan Bai, Yihang Lou, Xianzheng Ma 等AAAI 2023 · 被引用 145 次
