Universal Natural Language Processing with Limited Annotations: Try Few-shot Textual Entailment as a Start
Wenpeng Yin, Nazneen Fatema Rajani, Dragomir R. Radev, Richard Socher, Caiming Xiong
Abstract
A standard way to address different NLP problems is by first constructing a problem-specific dataset, then building a model to fit this dataset. To build the ultimate artificial intelligence, we desire a single machine that can handle diverse new problems, for which task-specific annotations are limited. We bring up textual entailment as a unified solver for such NLP problems. However, current research of textual entailment has not spilled much ink on the following questions: (i) How well does a pretrained textual entailment system generalize across domains with only a handful of domainspecific examples? and (ii) When is it worth transforming an NLP task into textual entailment? We argue that the transforming is unnecessary if we can obtain rich annotations for this task. Textual entailment really matters particularly when the target NLP task has insufficient annotations. Universal NLP 1 can be probably achieved through different routines. In this work, we introduce Universal Few-shot textual Entailment (UFO-ENTAIL). We demonstrate that this framework enables a pretrained entailment model to work well on new entailment domains in a few-shot setting, and show its effectiveness as a unified solver for several downstream NLP tasks such as question answering and coreference resolution when the end-task annotations are limited.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff18ea38-0cb8-43ce-8701-8a3184ba6b0bCited by top-tier papers17
- Generating Training Data with Language Models: Towards Zero-Shot Language UnderstandingYu Meng, Jiaxin Huang, Yu Zhang, Jiawei HanNeurIPS 2022 · 309 citations
- UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language ModelsTianbao Xie, Chen Henry Wu, Peng Shi, Ruiqi Zhong et al.EMNLP 2022 · 222 citations
- Differentiable Prompt Makes Pre-trained Language Models Better Few-shot LearnersNingyu Zhang, Luoqiu Li, Xiang Chen, Shumin Deng et al.ICLR 2022 · 205 citations
- CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLPQinyuan Ye, Bill Yuchen Lin, Xiang RenEMNLP 2021 · 103 citations
- Label Verbalization and Entailment for Effective Zero and Few-Shot Relation ExtractionOscar Sainz, Oier Lopez de Lacalle, Gorka Labaka, Ander Barrena et al.EMNLP 2021 · 94 citations
Related papers
- Entailment as Robust Self-LearnerJiaxin Ge, Hongyin Luo, Yoon Kim, James R. GlassACL 2023 · 2 citations
- UniSumm and SummZoo: Unified Model and Diverse Benchmark for Few-Shot SummarizationYulong Chen, Yang Liu, Ruochen Xu, Ziyi Yang et al.ACL 2023 · 7 citations
- Few-Shot Learning with Siamese Networks and Label TuningThomas Müller, Guillermo Pérez-Torró, Marc Franco-SalvadorACL 2022
- UniGen: Universal Domain Generalization for Sentiment Classification via Zero-shot Dataset GenerationJuhwan Choi, Yeonghwa Kim, Seunguk Yu, Jungmin Yun et al.EMNLP 2024 · 7 citations
- Single Domain Generalization for Few-Shot Counting via Universal Representation MatchingXianing Chen, Si Huo, Borui Jiang, Hailin Hu et al.CVPR 2025
