Label Semantic Aware Pre-training for Few-shot Text Classification
Aaron Mueller, Jason Krone, Salvatore Romeo, Saab Mansour, Elman Mansimov, Yi Zhang, Dan Roth
Abstract
In text classification tasks, useful information is encoded in the label names. Label semantic aware systems have leveraged this information for improved text classification performance during fine-tuning and prediction. However, use of label-semantics during pre-training has not been extensively explored. We therefore propose Label Semantic Aware Pre-training (LSAP) to improve the generalization and data efficiency of text classification systems. LSAP incorporates label semantics into pre-trained generative models (T5 in our case) by performing secondary pre-training on labeled sentences from a variety of domains. As domain-general pre-training requires large amounts of data, we develop a filtering and labeling pipeline to automatically create sentence-label pairs from unlabeled text. We perform experiments on intent (ATIS, Snips, TOPv2) and topic classification (AG News, Yahoo! Answers). LSAP obtains significant accuracy improvements over state-of-the-art models for few-shot text classification while maintaining performance comparable to state of the art in high-resource settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Zero-Shot Learners for Natural Language Understanding via a Unified Multiple Choice PerspectivePing Yang, Junjie Wang, Ruyi Gan, Xinyu Zhu et al.EMNLP 2022 · 10 citations
- Modeling Label Correlations for Ultra-Fine Entity Typing with Neural Pairwise Conditional Random FieldChengyue Jiang, Yong Jiang, Weiqi Wu, Pengjun Xie et al.EMNLP 2022 · 4 citations
- UniEX: An Effective and Efficient Framework for Unified Information Extraction via a Span-extractive PerspectiveYang Ping, Junyu Lu, Ruyi Gan, Junjie Wang et al.ACL 2023 · 4 citations
- Pre-training Intent-Aware Encoders for Zero- and Few-Shot Intent ClassificationMujeen Sung, James Gung, Elman Mansimov, Nikolaos Pappas et al.EMNLP 2023 · 1 citation
- Do LLMs Adhere to Label Definitions? Examining Their Receptivity to External Label DefinitionsSeyedali Mohammadi, Bhaskara Hanuma Vedula, Hemank Lamba, Edward Raff et al.EMNLP 2025
Builds on8
- Towards Scalable Multi-Domain Conversational Agents: The Schema-Guided Dialogue DatasetAbhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta et al.AAAI 2020 · 707 citations
- The Power of Scale for Parameter-Efficient Prompt TuningBrian Lester, Rami Al-Rfou, Noah ConstantEMNLP 2021 · 94 citations
- Low-Resource Domain Adaptation for Compositional Task-Oriented Semantic ParsingXilun Chen, Asish Ghoshal, Yashar Mehdad, Luke Zettlemoyer et al.EMNLP 2020 · 66 citations
- Discriminative Nearest Neighbor Few-Shot Intent Detection by Transferring Natural Language InferenceJian-Guo Zhang, Kazuma Hashimoto, Wenhao Liu, Chien-Sheng Wu et al.EMNLP 2020 · 65 citations
- Template Guided Text Generation for Task-Oriented DialogueMihir Kale, Abhinav RastogiEMNLP 2020 · 56 citations
Related papers
- Augmented Natural Language for Generative Sequence LabelingBen Athiwaratkun, Cícero Nogueira dos Santos, Jason Krone, Bing XiangEMNLP 2020 · 54 citations
- The Benefits of Label-Description Training for Zero-Shot Text ClassificationLingyu Gao, Debanjan Ghosh, Kevin GimpelEMNLP 2023 · 6 citations
- Rethinking the Effect of Uninformative Class Name in Prompt LearningFengmao Lv, Changru Nie, Jianyang Zhang, Guowu Yang et al.ACM MM 2024 · 1 citation
- CINS: Comprehensive Instruction for Few-Shot Learning in Task-Oriented Dialog SystemsFei Mi, Yasheng Wang, Yitong LiAAAI 2022 · 46 citations
- Few-Shot Learning with Siamese Networks and Label TuningThomas Müller, Guillermo Pérez-Torró, Marc Franco-SalvadorACL 2022
