Less Is Better: Sparse Instance Learning for Cross-Domain Few-Shot Object Detection
Yali Huang, Jie Mei, Ziyi Wu, Yiming Yang, Hongru Zhao, Mingyuan Jiu, Hichem Sahbi
Abstract
Cross-Domain Few-Shot Object Detection (CD-FSOD) is an extremely challenging task due to the inherent data scarcity and substantial domain shift between the source and target domains. Existing methods often suffer from overfitting and noisy feature representations, which hinder the construction of discriminative class prototypes in the target domain. In this paper, we propose a novel framework with sparse instance learning (SI-ViTO) for CD-FSOD, which leverages instance sparsity to achieve a better detection with less representation. SI-ViTO adopts a dual-stage sparsity module, consisting of instance feature sparsity not only on the few-shot support images but also on the query images. This dual sparsity enables the model to effectively preserve salient foreground semantics and simultaneously to filter out redundant or noisy information. Furthermore, a new prototype calibration strategy is also used to dynamically refine the class prototypes with query instances to accelerate prototype adaptation. Extensive experimental results on CD-FSOD benchmarks show that SI-ViTO outperforms the state-of-the-art methods, demonstrating that less discriminative representations yield better crossdomain few-shot object detection performance than more abundant ones.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 74283c65-3d8b-4181-af8d-07f3b45efa3aBuilds on15
- AdaptFormer: Adapting Vision Transformers for Scalable Visual RecognitionShoufa Chen, Chongjian Ge, Zhan Tong, Jiangliu Wang et al.NeurIPS 2022 · 1,291 citations
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell et al.ICML 2020 · 723 citations
- Meta R-CNN: Towards General Solver for Instance-Level Low-Shot LearningXiaopeng Yan, Ziliang Chen, Anni Xu, Xiaoxi Wang et al.ICCV 2019 · 590 citations
- DeFRCN: Decoupled Faster R-CNN for Few-Shot Object DetectionLimeng Qiao, Yuxuan Zhao, Zhiyuan Li, Xi Qiu et al.ICCV 2021 · 298 citations
- Sparse DETR: Efficient End-to-End Object Detection with Learnable SparsityByungseok Roh, Jaewoong Shin, Wuhyun Shin, Saehoon KimICLR 2022 · 256 citations
Related papers
- StyleProto: Style-Augmented Prototype Learning for Cross-Domain Few-Shot Object DetectionXi Yang, Quantao XieAAAI 2026
- AsyFOD: An Asymmetric Adaptation Paradigm for Few-Shot Domain Adaptive Object DetectionYipeng Gao, Kun-Yu Lin, Junkai Yan, Yaowei Wang et al.CVPR 2023
- A Closer Look at the CLS Token for Cross-Domain Few-Shot LearningYixiong Zou, Shuai Yi, Yuhua Li, Ruixuan LiNeurIPS 2024 · 40 citations
- Meta Faster R-CNN: Towards Accurate Few-Shot Object Detection with Attentive Feature AlignmentGuangxing Han, Shiyuan Huang, Jiawei Ma, Yicheng He et al.AAAI 2022 · 227 citations
- Remedying Target-Domain Astigmatism for Cross-Domain Few-Shot Object DetectionYongwei Jiang, Yixiong Zou, Yuhua Li, Ruixuan LiCVPR 2026 · 3 citations
