An Analysis of Natural Language Inference Benchmarks through the Lens of Negation
Md Mosharaf Hossain, Venelin Kovatchev, Pranoy Dutta, Tiffany Kao, Elizabeth Wei, Eduardo Blanco
Abstract
Negation is underrepresented in existing natural language inference benchmarks. Additionally, one can often ignore the few negations in existing benchmarks and still make the right inference judgments. In this paper, we present a new benchmark for natural language inference in which negation plays an important role. We also show that state-of-the-art transformers struggle making inference judgments with the new pairs. Original pair New pair w/ negation RTE T: Tropical Storm Debby is blamed for several deaths across the Caribbean. T neg : Tropical Storm Debby is not blamed for several deaths across the Caribbean. H: A tropical storm has caused loss of life. H neg : A tropical storm has not caused loss of life.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- Evaluating Large Language Models at Evaluating Instruction FollowingZhiyuan Zeng, Jiatong Yu, Tianyu Gao, Yu Meng et al.ICLR 2024 · 299 citations
- Consistency Analysis of ChatGPTMyeongjun Jang, Thomas LukasiewiczEMNLP 2023 · 55 citations
- Learning Deductive Reasoning from Synthetic Corpus based on Formal LogicTerufumi Morishita, Gaku Morio, Atsuki Yamaguchi, Yasuhiro SogawaICML 2023 · 45 citations
- Diagnosing the First-Order Logical Reasoning Ability Through LogicNLIJidong Tian, Yitian Li, Wenqing Chen, Liqiang Xiao et al.EMNLP 2021 · 21 citations
- Multi-VALUE: A Framework for Cross-Dialectal English NLPCaleb Ziems, William Barr Held, Jingfeng Yang, Jwala Dhamala et al.ACL 2023 · 15 citations
Builds on2
- Beyond Accuracy: Behavioral Testing of NLP Models with CheckListMarco Túlio Ribeiro, Tongshuang Wu, Carlos Guestrin, Sameer SinghACL 2020 · 51 citations
- Predicting the Focus of Negation: Model and Error AnalysisMd Mosharaf Hossain, Kathleen E. Hamilton, Alexis Palmer, Eduardo BlancoACL 2020 · 8 citations
Related papers
- Vision-Language Models Do Not Understand NegationKumail Alhamoud, Shaden Alshammari, Yonglong Tian, Guohao Li et al.CVPR 2025
- Leveraging Affirmative Interpretations from Negation Improves Natural Language UnderstandingMd Mosharaf Hossain, Eduardo BlancoEMNLP 2022 · 4 citations
- RobustLR: A Diagnostic Benchmark for Evaluating Logical Robustness of Deductive ReasonersSoumya Sanyal, Zeyi Liao, Xiang RenEMNLP 2022 · 6 citations
- Enhancing Natural Language Inference Using New and Expanded Training Data Sets and New Learning ModelsArindam Mitra, Ishan Shrivastava, Chitta BaralAAAI 2020 · 10 citations
- Learn to Understand Negation in Video RetrievalZiyue Wang, Aozhu Chen, Fan Hu, Xirong LiACM MM 2022 · 12 citations
