Label Hallucination for Few-Shot Classification
Yiren Jian, Lorenzo Torresani
Abstract
Few-shot classification requires adapting knowledge learned from a large annotated base dataset to recognize novel unseen classes, each represented by few labeled examples. In such a scenario, pretraining a network with high capacity on the large dataset and then finetuning it on the few examples causes severe overfitting. At the same time, training a simple linear classifier on top of frozen'' features learned from the large labeled dataset fails to adapt the model to the properties of the novel classes, effectively inducing underfitting. In this paper we propose an alternative approach to both of these two popular strategies. First, our method pseudo-labels the entire large dataset using the linear classifier trained on the novel classes. This effectively hallucinates'' the novel classes in the large dataset, despite the novel categories not being present in the base database (novel and base classes are disjoint). Then, it finetunes the entire model with a distillation loss on the pseudo-labeled base examples, in addition to the standard cross-entropy loss on the novel dataset. This step effectively trains the network to recognize contextual and appearance cues that are useful for the novel-category recognition but using the entire large-scale base dataset and thus overcoming the inherent data-scarcity problem of few-shot learning. Despite the simplicity of the approach, we show that that our method outperforms the state-of-the-art on four well-established few-shot classification benchmarks. The code is available at https://github.com/yiren-jian/LabelHalluc.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6a855366-f736-4214-80eb-1309240ff4adCited by top-tier papers8
- Generating Images of Rare Concepts Using Pre-trained Diffusion ModelsDvir Samuel, Rami Ben-Ari, Simon Raviv, Nir Darshan et al.AAAI 2024 · 82 citations
- Norm-guided latent space exploration for text-to-image generationDvir Samuel, Rami Ben-Ari, Nir Darshan, Haggai Maron et al.NeurIPS 2023 · 49 citations
- FeLMi : Few shot Learning with hard MixupAniket Roy, Anshul Shah, Ketul Shah, Prithviraj Dhar et al.NeurIPS 2022 · 41 citations
- Boosting Few-Shot Learning via Attentive Feature RegularizationXingyu Zhu, Shuo Wang, Jinda Lu, Yanbin Hao et al.AAAI 2024 · 30 citations
- RankDNN: Learning to Rank for Few-Shot LearningQianyu Guo, Haotong Gong, Xujun Wei, Yanwei Fu et al.AAAI 2023 · 27 citations
Builds on19
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- A Baseline for Few-Shot Image ClassificationGuneet Singh Dhillon, Pratik Chaudhari, Avinash Ravichandran, Stefano SoattoICLR 2020 · 640 citations
- Theoretical Analysis of Self-Training with Deep Networks on Unlabeled DataColin Wei, Kendrick Shen, Yining Chen, Tengyu MaICLR 2021 · 261 citations
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin et al.ICLR 2020 · 221 citations
Related papers
- Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification ReframingHan Liu, Siyang Zhao, Xiaotong Zhang, Feng Zhang et al.AAAI 2024 · 7 citations
- On the Importance of Distractors for Few-Shot ClassificationRajshekhar Das, Yu-Xiong Wang, José M. F. MouraICCV 2021 · 36 citations
- Incremental-DETR: Incremental Few-Shot Object Detection via Self-Supervised LearningNa Dong, Yongqiang Zhang, Mingli Ding, Gim Hee LeeAAAI 2023 · 54 citations
- An Embarrassingly Simple Approach to Semi-Supervised Few-Shot LearningXiu-Shen Wei, He-Yang Xu, Faen Zhang, Yuxin Peng et al.NeurIPS 2022 · 24 citations
- Instance Credibility Inference for Few-Shot LearningYikai Wang, Chengming Xu, Chen Liu, Li Zhang et al.CVPR 2020
