Looking at the Overlooked: An Analysis on the Word-Overlap Bias in Natural Language Inference
Sara Rajaee, Yadollah Yaghoobzadeh, Mohammad Taher Pilehvar
摘要
It has been shown that NLI models are usually biased with respect to the word-overlap between premise and hypothesis; they take this feature as a primary cue for predicting the entailment label. In this paper, we focus on an overlooked aspect of the overlap bias in NLI models: the reverse word-overlap bias. Our experimental results demonstrate that current NLI models are highly biased towards the non-entailment label on instances with low overlap, and the existing debiasing methods, which are reportedly successful on existing challenge datasets, are generally ineffective in addressing this category of bias. We investigate the reasons for the emergence of the overlap bias and the role of minority examples in its mitigation. For the former, we find that the word-overlap bias does not stem from pre-training, and for the latter, we observe that in contrast to the accepted assumption, eliminating minority examples does not affect the generalizability of debiasing methods with respect to the overlap bias. All the code and relevant data are available at: https: //github.com/sara-rajaee/reverse_bias Overlap Sample Label Full (1.0) P: A little kid in blue is sledding down a snowy hill. H: A little kid in blue sledding. Entailment P: The young lady is giving the old man a hug. H: The young man is giving the old man a hug. Non-Entailment 12 13 = 0.923 P: A woman in a blue shirt and green hat looks up at the camera. H: A woman wearing a blue shirt and green hat looks at the camera Entailment 11 12 = 0.917 P: Two men in wheelchairs are reaching in the air for a basketball. H: Two women in wheelchairs are reaching in the air for a basketball. Non-Entailment 1 14 = 0.071 P: Several young people sit at a table playing poker. H: Youthful Human beings are gathered around a flat surface to play a card game. Entailment
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse and MoreOuail Kitouni, Niklas Nolte, Adina Williams, Michael Rabbat 等NeurIPS 2024 · 被引用 29 次
- Improving the robustness of NLI models with minimax trainingMichalis Korakakis, Andreas VlachosACL 2023 · 被引用 4 次
- RepMatch: Quantifying Cross-Instance Similarities in Representation SpaceMohammad Modarres, Sina Abbasi, Mohammad Taher PilehvarEMNLP 2024
它引用的顶会 Paper10
- End-to-End Bias Mitigation by Modelling Biases in CorporaRabeeh Karimi Mahabadi, Yonatan Belinkov, James HendersonACL 2020 · 被引用 136 次
- Learning from others' mistakes: Avoiding dataset biases without modeling themVictor Sanh, Thomas Wolf, Yonatan Belinkov, Alexander M. RushICLR 2021 · 被引用 123 次
- Generating Data to Mitigate Spurious Correlations in Natural Language Inference DatasetsYuxiang Wu, Matt Gardner, Pontus Stenetorp, Pradeep DasigiACL 2022 · 被引用 74 次
- Competency Problems: On Finding and Removing Artifacts in Language DataMatt Gardner, William Merrill, Jesse Dodge, Matthew E. Peters 等EMNLP 2021 · 被引用 72 次
- Towards Robustifying NLI Models Against Lexical Dataset BiasesXiang Zhou, Mohit BansalACL 2020 · 被引用 36 次
相关 Paper
- Mitigating Spurious Correlation in Natural Language Understanding with Counterfactual InferenceCan Udomcharoenchaikit, Wuttikorn Ponwitayarat, Patomporn Payoungkhamdee, Kanruethai Masuk 等EMNLP 2022 · 被引用 10 次
- Supervising Model Attention with Human Explanations for Robust Natural Language InferenceJoe Stacey, Yonatan Belinkov, Marek ReiAAAI 2022 · 被引用 52 次
- Avoiding the Hypothesis-Only Bias in Natural Language Inference via Ensemble Adversarial TrainingJoe Stacey, Pasquale Minervini, Haim Dubossarsky, Sebastian Riedel 等EMNLP 2020 · 被引用 6 次
- IRM - when it works and when it doesn't: A test case of natural language inferenceYana Dranker, He He, Yonatan BelinkovNeurIPS 2021 · 被引用 22 次
- MABEL: Attenuating Gender Bias using Textual Entailment DataJacqueline He, Mengzhou Xia, Christiane Fellbaum, Danqi ChenEMNLP 2022 · 被引用 15 次
