Few-Shot Learning with Siamese Networks and Label Tuning
Thomas Müller, Guillermo Pérez-Torró, Marc Franco-Salvador
Abstract
We study the problem of building text classifiers with little or no training data, commonly known as zero and few-shot text classification. In recent years, an approach based on neural textual entailment models has been found to give strong results on a diverse range of tasks. In this work, we show that with proper pre-training, Siamese Networks that embed texts and labels offer a competitive alternative. These models allow for a large reduction in inference cost: constant in the number of labels rather than linear. Furthermore, we introduce label tuning, a simple and computationally efficient approach that allows to adapt the models in a few-shot setup by only changing the label embeddings. While giving lower performance than model fine-tuning, this approach has the architectural advantage that a single encoder can be shared by many different tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6fc73ecf-20a8-4acc-bf72-a01ea97512fdCited by top-tier papers2
- MoEEdit: Efficient and Routing-Stable Knowledge Editing for Mixture-of-Experts LLMsYupu Gu, Rongzhe Wei, Andy Zhu, Pan LiICLR 2026 · 4 citations
- Was My Data Used for Training? Membership Inference in Open-Source LLMs via Neural ActivationsXue Tan, Hao Luan, Mingyu Luo, Zhuyang Yu et al.NDSS 2026 · 2 citations
Builds on6
- MPNet: Masked and Permuted Pre-training for Language UnderstandingKaitao Song, Xu Tan, Tao Qin, Jianfeng Lu et al.NeurIPS 2020 · 1,957 citations
- Adversarial NLI: A New Benchmark for Natural Language UnderstandingYixin Nie, Adina Williams, Emily Dinan, Mohit Bansal et al.ACL 2020 · 602 citations
- True Few-Shot Learning with Language ModelsEthan Perez, Douwe Kiela, Kyunghyun ChoNeurIPS 2021 · 547 citations
- Universal Natural Language Processing with Limited Annotations: Try Few-shot Textual Entailment as a StartWenpeng Yin, Nazneen Fatema Rajani, Dragomir R. Radev, Richard Socher et al.EMNLP 2020 · 57 citations
- WARP: Word-level Adversarial ReProgrammingKaren Hambardzumyan, Hrant Khachatrian, Jonathan MayACL 2021
Related papers
- Label Verbalization and Entailment for Effective Zero and Few-Shot Relation ExtractionOscar Sainz, Oier Lopez de Lacalle, Gorka Labaka, Ander Barrena et al.EMNLP 2021 · 94 citations
- Semantic matching for text classification with complex class descriptionsBrian de Silva, Kuan-Wen Huang, Gwang Lee, Karen Hovsepian et al.EMNLP 2023 · 1 citation
- PromptBoosting: Black-Box Text Classification with Ten Forward PassesBairu Hou, Joe O'Connor, Jacob Andreas, Shiyu Chang et al.ICML 2023 · 56 citations
- Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification ReframingHan Liu, Siyang Zhao, Xiaotong Zhang, Feng Zhang et al.AAAI 2024 · 7 citations
- The Benefits of Label-Description Training for Zero-Shot Text ClassificationLingyu Gao, Debanjan Ghosh, Kevin GimpelEMNLP 2023 · 6 citations
