CDFSL-V: Cross-Domain Few-Shot Learning for Videos
Sarinda Samarasinghe, Mamshad Nayeem Rizve, Navid Kardan, Mubarak Shah
Abstract
Few-shot video action recognition is an effective approach to recognizing new categories with only a few labeled examples, thereby reducing the challenges associated with collecting and annotating large-scale video datasets. Existing methods in video action recognition rely on large labeled datasets from the same domain. However, this setup is not realistic as novel categories may come from different data domains that may have different spatial and temporal characteristics. This dissimilarity between the source and target domains can pose a significant challenge, rendering traditional few-shot action recognition techniques ineffective. To address this issue, in this work, we propose a novel cross-domain few-shot video action recognition method that leverages self-supervised learning and curriculum learning to balance the information from the source and target domains. To be particular, our method employs a masked autoencoder-based self-supervised training objective to learn from both source and target data in a self-supervised manner. Then a progressive curriculum balances learning the discriminative information from the source dataset with the generic information learned from the target domain. Initially, our curriculum utilizes supervised learning to learn class discriminative features from the source data. As the training progresses, we transition to learning target-domain-specific features. We propose a progressive curriculum to encourage the emergence of rich features in the target domain based on class discriminative supervised features in the source domain. We evaluate our method on several challenging benchmark datasets and demonstrate that our approach outperforms existing cross-domain few-shot learning techniques. Our code is available at https://github.com/Sarinda251/CDFSL-V
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 69fa85d2-a725-4f16-8e6c-290195e0c420Cited by top-tier papers5
- Learning Causal Domain-Invariant Temporal Dynamics for Few-Shot Action RecognitionYuke Li, Guangyi Chen, Ben Abramowitz, Stefano Anzellotti et al.ICML 2024 · 3 citations
- TAMT: Temporal-Aware Model Tuning for Cross-Domain Few-Shot Action RecognitionYilong Wang, Zilin Gao, Qilong Wang, Zhaofeng Chen et al.CVPR 2025
- Temporal Alignment-Free Video Matching for Few-shot Action RecognitionSuBeen Lee, WonJun Moon, Hyun Seok Seong, Jae-Pil HeoCVPR 2025
- Cross-Domain Few-Shot Segmentation via Multi-view Progressive AdaptationJiahao Nie, Guanqiao Fu, Wenbin An, Yap-Peng Tan et al.CVPR 2026
- Harnessing Spectrum Video for Subject-Level Few-Shot and Cross-Montage EEG GeneralizationWei Wang, Fang He, Yifan Li, Wanying Qu et al.ICML 2026
Builds on10
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-TrainingZhan Tong, Yibing Song, Jue Wang, Limin WangNeurIPS 2022 · 2,336 citations
- Open-World Semi-Supervised LearningKaidi Cao, Maria Brbic, Jure LeskovecICLR 2022 · 246 citations
- Spatio-temporal Relation Modeling for Few-shot Action RecognitionAnirudh Thatipelli, Sanath Narayan, Salman Khan, Rao Muhammad Anwer et al.CVPR 2022 · 144 citations
- Self-training For Few-shot Transfer Across Extreme Task DifferencesCheng Perng Phoo, Bharath HariharanICLR 2021 · 131 citations
Related papers
- Reconstruction Target Matters in Masked Image Modeling for Cross-Domain Few-Shot LearningRan Ma, Yixiong Zou, Yuhua Li, Ruixuan LiAAAI 2025 · 2 citations
- On the Importance of Spatial Relations for Few-shot Action RecognitionYilun Zhang, Yuqian Fu, Xingjun Ma, Lizhe Qi et al.ACM MM 2023 · 20 citations
- Understanding Cross-Domain Few-Shot Learning Based on Domain Similarity and Few-Shot DifficultyJaehoon Oh, Sungnyun Kim, Namgyu Ho, Jin-Hwa Kim et al.NeurIPS 2022 · 69 citations
- Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-TrainingArun V. Reddy, William Paul, Corban Rivera, Ketul Shah et al.CVPR 2024 · 3 citations
- Dynamic Distillation Network for Cross-Domain Few-Shot Recognition with Unlabeled DataAshraful Islam, Chun-Fu (Richard) Chen, Rameswar Panda, Leonid Karlinsky et al.NeurIPS 2021 · 106 citations
