Label Semantic Aware Pre-training for Few-shot Text Classification
Aaron Mueller, Jason Krone, Salvatore Romeo, Saab Mansour, Elman Mansimov, Yi Zhang, Dan Roth
摘要
In text classification tasks, useful information is encoded in the label names. Label semantic aware systems have leveraged this information for improved text classification performance during fine-tuning and prediction. However, use of label-semantics during pre-training has not been extensively explored. We therefore propose Label Semantic Aware Pre-training (LSAP) to improve the generalization and data efficiency of text classification systems. LSAP incorporates label semantics into pre-trained generative models (T5 in our case) by performing secondary pre-training on labeled sentences from a variety of domains. As domain-general pre-training requires large amounts of data, we develop a filtering and labeling pipeline to automatically create sentence-label pairs from unlabeled text. We perform experiments on intent (ATIS, Snips, TOPv2) and topic classification (AG News, Yahoo! Answers). LSAP obtains significant accuracy improvements over state-of-the-art models for few-shot text classification while maintaining performance comparable to state of the art in high-resource settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Zero-Shot Learners for Natural Language Understanding via a Unified Multiple Choice PerspectivePing Yang, Junjie Wang, Ruyi Gan, Xinyu Zhu 等EMNLP 2022 · 被引用 10 次
- Modeling Label Correlations for Ultra-Fine Entity Typing with Neural Pairwise Conditional Random FieldChengyue Jiang, Yong Jiang, Weiqi Wu, Pengjun Xie 等EMNLP 2022 · 被引用 4 次
- UniEX: An Effective and Efficient Framework for Unified Information Extraction via a Span-extractive PerspectiveYang Ping, Junyu Lu, Ruyi Gan, Junjie Wang 等ACL 2023 · 被引用 4 次
- Pre-training Intent-Aware Encoders for Zero- and Few-Shot Intent ClassificationMujeen Sung, James Gung, Elman Mansimov, Nikolaos Pappas 等EMNLP 2023 · 被引用 1 次
- Do LLMs Adhere to Label Definitions? Examining Their Receptivity to External Label DefinitionsSeyedali Mohammadi, Bhaskara Hanuma Vedula, Hemank Lamba, Edward Raff 等EMNLP 2025
它引用的顶会 Paper8
- Towards Scalable Multi-Domain Conversational Agents: The Schema-Guided Dialogue DatasetAbhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta 等AAAI 2020 · 被引用 707 次
- The Power of Scale for Parameter-Efficient Prompt TuningBrian Lester, Rami Al-Rfou, Noah ConstantEMNLP 2021 · 被引用 94 次
- Low-Resource Domain Adaptation for Compositional Task-Oriented Semantic ParsingXilun Chen, Asish Ghoshal, Yashar Mehdad, Luke Zettlemoyer 等EMNLP 2020 · 被引用 66 次
- Discriminative Nearest Neighbor Few-Shot Intent Detection by Transferring Natural Language InferenceJian-Guo Zhang, Kazuma Hashimoto, Wenhao Liu, Chien-Sheng Wu 等EMNLP 2020 · 被引用 65 次
- Template Guided Text Generation for Task-Oriented DialogueMihir Kale, Abhinav RastogiEMNLP 2020 · 被引用 56 次
相关 Paper
- Augmented Natural Language for Generative Sequence LabelingBen Athiwaratkun, Cícero Nogueira dos Santos, Jason Krone, Bing XiangEMNLP 2020 · 被引用 54 次
- The Benefits of Label-Description Training for Zero-Shot Text ClassificationLingyu Gao, Debanjan Ghosh, Kevin GimpelEMNLP 2023 · 被引用 6 次
- Rethinking the Effect of Uninformative Class Name in Prompt LearningFengmao Lv, Changru Nie, Jianyang Zhang, Guowu Yang 等ACM MM 2024 · 被引用 1 次
- CINS: Comprehensive Instruction for Few-Shot Learning in Task-Oriented Dialog SystemsFei Mi, Yasheng Wang, Yitong LiAAAI 2022 · 被引用 46 次
- Few-Shot Learning with Siamese Networks and Label TuningThomas Müller, Guillermo Pérez-Torró, Marc Franco-SalvadorACL 2022
