Rethinking Generalization in Few-Shot Classification
Markus Hiller, Rongkai Ma, Mehrtash Harandi, Tom Drummond
Abstract
Single image-level annotations only correctly describe an often small subset of an image's content, particularly when complex real-world scenes are depicted. While this might be acceptable in many classification scenarios, it poses a significant challenge for applications where the set of classes differs significantly between training and test time. In this paper, we take a closer look at the implications in the context of . Splitting the input samples into patches and encoding these via the help of Vision Transformers allows us to establish semantic correspondences between local regions across images and independent of their respective class. The most informative patch embeddings for the task at hand are then determined as a function of the support set via online optimization at inference time, additionally providing visual interpretability of `' in the image. We build on recent advances in unsupervised training of networks via masked image modelling to overcome the lack of fine-grained labels and learn the more general statistical structure of the data while avoiding negative image-level annotation influence, supervision collapse. Experimental results show the competitiveness of our approach, achieving new state-of-the-art results on four popular few-shot classification benchmarks for -shot and -shot scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 92566a08-68dc-4598-a75b-210fa80dadbbCited by top-tier papers18
- Class-Aware Patch Embedding Adaptation for Few-Shot Image ClassificationFusheng Hao, Fengxiang He, Liu Liu, Fuxiang Wu et al.ICCV 2023 · 56 citations
- Strong Baselines for Parameter-Efficient Few-Shot Fine-TuningSamyadeep Basu, Shell Xu Hu, Daniela Massiceti, Soheil FeiziAAAI 2024 · 54 citations
- Simple Semantic-Aided Few-Shot LearningHai Zhang, Junzhe Xu, Shanlin Jiang, Zhenan HeCVPR 2024 · 33 citations
- Focus Your Attention when Few-Shot ClassificationHaoqing Wang, Shibo Jie, Zhihong DengNeurIPS 2023 · 16 citations
- Envisioning Class Entity Reasoning by Large Language Models for Few-shot LearningMushui Liu, Fangtai Wu, Bozheng Li, Ziqian Lu et al.AAAI 2025 · 15 citations
Builds on36
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
Related papers
- Supervised Masked Knowledge Distillation for Few-Shot TransformersHan Lin, Guangxing Han, Jiawei Ma, Shiyuan Huang et al.CVPR 2023
- CAD: Co-Adapting Discriminative Features for Improved Few-Shot ClassificationPhilip Chikontwe, Soopil Kim, Sang Hyun ParkCVPR 2022 · 46 citations
- Distilling Self-Supervised Vision Transformers for Weakly-Supervised Few-Shot Classification & SegmentationDahyun Kang, Piotr Koniusz, Minsu Cho, Naila MurrayCVPR 2023
- Label-Efficient Few-Shot Semantic Segmentation with Unsupervised Meta-TrainingJianwu Li, Kaiyue Shi, Guo-Sen Xie, Xiaofeng Liu et al.AAAI 2024 · 15 citations
- Pushing the Limits of Simple Pipelines for Few-Shot Learning: External Data and Fine-Tuning Make a DifferenceShell Xu Hu, Da Li, Jan Stühmer, Minyoung Kim et al.CVPR 2022 · 161 citations
