What does the Failure to Reason with "Respectively" in Zero/Few-Shot Settings Tell Us about Language Models?
Ruixiang Cui, Seolhwa Lee, Daniel Hershcovich, Anders Søgaard
摘要
Humans can effortlessly understand the coordinate structure of sentences such as "Niels Bohr and Kurt Cobain were born in Copenhagen and Seattle, respectively". In the context of natural language inference (NLI), we examine how language models (LMs) reason with respective readings (Gawron and Kehler, 2004) from two perspectives: syntactic-semantic and commonsense-world knowledge. We propose a controlled synthetic dataset WikiResNLI and a naturally occurring dataset NatResNLI to encompass various explicit and implicit realizations of "respectively". We show that finetuned NLI models struggle with understanding such readings without explicit supervision. While few-shot learning is easy in the presence of explicit cues, longer training is required when the reading is evoked implicitly, leaving models to rely on common sense inferences. Furthermore, our fine-grained analysis indicates models fail to generalize across different constructions. To conclude, we demonstrate that LMs still lag behind humans in generalizing to the long tail of linguistic constructions. Denotation Natural Language Example Premise: w1 and w3 p w2 and w4 , respectively. Emiliano Zapata and Gerhart Münch died in Morelos and Michoacán , respectively Hypotheses: Entailment (1), 1S1O w1 p w2 . Emiliano Zapata died in Morelos . Entailment (2), 1S1O w3 p w4 . Gerhart Münch died in Michoacán . Contradiction (1), 1S1O w1 p w4 . Emiliano Zapata died in Michoacán . Contradiction (2), 1S1O w3 p w2 . Gerhart Münch died in Morelos . Contradiction (3), 1S2O w1 p w2 and w4 . Emiliano Zapata died in Morelos and Michoacán . Contradiction (4), 1S2O w3 p w2 and w4 . Gerhart Münch died in Morelos and Michoacán . Contradiction (5), 2S1O w1 and w3 p w2 . Emiliano Zapata and Gerhart Münch died in Morelos . Contradiction (6), 2S1O w1 and w3 p w4 . Emiliano Zapata and Gerhart Münch died in Michoacán .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- Adversarial NLI: A New Benchmark for Natural Language UnderstandingYixin Nie, Adina Williams, Emily Dinan, Mohit Bansal 等ACL 2020 · 被引用 602 次
- Large Language Models Can Self-ImproveJiaxin Huang, Shixiang Gu, Le Hou, Yuexin Wu 等EMNLP 2023 · 被引用 184 次
相关 Paper
- This is not a Dataset: A Large Negation Benchmark to Challenge Large Language ModelsIker García-Ferrero, Begoña Altuna, Javier Álvez, Itziar Gonzalez-Dios 等EMNLP 2023 · 被引用 8 次
- Entailed Between the Lines: Incorporating Implication into NLIShreya Havaldar, Hamidreza Alvari, John Palowitch, Mohammad Javad Hosseini 等ACL 2025
- Natural Language Inference in Context - Investigating Contextual Reasoning over Long TextsHanmeng Liu, Leyang Cui, Jian Liu, Yue ZhangAAAI 2021 · 被引用 57 次
- IMPLI: Investigating NLI Models' Performance on Figurative LanguageKevin Stowe, Prasetya Ajie Utama, Iryna GurevychACL 2022 · 被引用 52 次
- Are Natural Language Inference Models IMPPRESsive? Learning IMPlicature and PRESuppositionPaloma Jeretic, Alex Warstadt, Suvrat Bhooshan, Adina WilliamsACL 2020 · 被引用 2 次
