ProD: Prompting-to-disentangle Domain Knowledge for Cross-domain Few-shot Image Classification
Tianyi Ma, Yifan Sun, Zongxin Yang, Yi Yang
摘要
This paper considers few-shot image classification under the cross-domain scenario, where the train-to-test domain gap compromises classification accuracy. To mitigate the domain gap, we propose a prompting-to-disentangle (ProD) method through a novel exploration with the prompting mechanism. ProD adopts the popular multi-domain training scheme and extracts the backbone feature with a standard Convolutional Neural Network. Based on these two common practices, the key point of ProD is using the prompting mechanism in the transformer to disentangle the domain-general (DG) and domain-specific (DS) knowledge from the backbone feature. Specifically, ProD concatenates a DG and a DS prompt to the backbone feature and feeds them into a lightweight transformer. The DG prompt is learnable and shared by all the training domains, while the DS prompt is generated from the domain-of-interest on the fly. As a result, the transformer outputs DG and DS features in parallel with the two prompts, yielding the disentangling effect. We show that: 1) Simply sharing a single DG prompt for all the training domains already improves generalization towards the novel test domain. 2) The cross-domain generalization can be further reinforced by making the DG prompt neutral towards the training domains. 3) When inference, the DS prompt is generated from the support samples and can capture test domain knowledge through the prompting mechanism. Combining all three benefits, ProD significantly improves cross-domain few-shot classification. For instance, on CUB, ProD improves the 5-way 5-shot accuracy from 73.56% (baseline) to 79.19%, setting a new state of the art.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Amend to Alignment: Decoupled Prompt Tuning for Mitigating Spurious Correlation in Vision-Language ModelsJie Zhang, Xiaosong Ma, Song Guo, Peng Li 等ICML 2024 · 被引用 10 次
- Random Registers for Cross-Domain Few-Shot LearningShuai Yi, Yixiong Zou, Yuhua Li, Ruixuan LiICML 2025
- One is Plenty: A Polymorphic Feature Interpreter for Immutable Heterogeneous Collaborative PerceptionYuchen Xia, Quan Yuan, Guiyang Luo, Xiaoyuan Fu 等CVPR 2025
- Adapt Before Comparison: A New Perspective on Cross-Domain Few-Shot SegmentationJonas HerzogCVPR 2024
它引用的顶会 Paper12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Cross-Domain Few-Shot Classification via Learned Feature-Wise TransformationHung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang, Ming-Hsuan YangICLR 2020 · 被引用 467 次
- Model Adaptation: Historical Contrastive Learning for Unsupervised Domain Adaptation without Source DataJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuNeurIPS 2021 · 被引用 301 次
- Invariant Risk Minimization GamesKartik Ahuja, Karthikeyan Shanmugam, Kush R. Varshney, Amit DhurandharICML 2020 · 被引用 289 次
相关 Paper
- Task-Adaptive Prompted Transformer for Cross-Domain Few-Shot LearningJiamin Wu, Xin Liu, Xiaotian Yin, Tianzhu Zhang 等AAAI 2024 · 被引用 14 次
- Meta-FDMixup: Cross-Domain Few-Shot Learning Guided by Labeled Target DataYuqian Fu, Yanwei Fu, Yu-Gang JiangACM MM 2021 · 被引用 85 次
- Cross-Level Distillation and Feature Denoising for Cross-Domain Few-Shot ClassificationHao Zheng, Runqi Wang, Jianzhuang Liu, Asako KanezakiICLR 2023 · 被引用 3 次
- Mind the Gap Between Prototypes and Images in Cross-domain FinetuningHongduan Tian, Feng Liu, Zhanke Zhou, Tongliang Liu 等NeurIPS 2024 · 被引用 4 次
- Switch to Generalize: Domain-Switch Learning for Cross-Domain Few-Shot ClassificationZhengdong Hu, Yifan Sun, Yi YangICLR 2022 · 被引用 22 次
