Towards Robustifying NLI Models Against Lexical Dataset Biases
Xiang Zhou, Mohit Bansal
摘要
While deep learning models are making fast progress on the task of Natural Language Inference, recent studies have also shown that these models achieve high accuracy by exploiting several dataset biases, and without deep understanding of the language semantics. Using contradiction-word bias and word-overlapping bias as our two bias examples, this paper explores both data-level and model-level debiasing methods to robustify models against lexical dataset biases. First, we debias the dataset through data augmentation and enhancement, but show that the model bias cannot be fully removed via this method. Next, we also compare two ways of directly debiasing the model without knowing what the dataset biases are in advance. The first approach aims to remove the label bias at the embedding level. The second approach employs a bag-of-words sub-model to capture the features that are likely to exploit the bias and prevents the original model from learning these biased features by forcing orthogonality between these two sub-models. We performed evaluations on new balanced datasets extracted from the original MNLI dataset as well as the NLI stress tests, and show that the orthogonality approach is better at debiasing the model while maintaining competitive overall accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Generating Data to Mitigate Spurious Correlations in Natural Language Inference DatasetsYuxiang Wu, Matt Gardner, Pontus Stenetorp, Pradeep DasigiACL 2022 · 被引用 74 次
- Looking at the Overlooked: An Analysis on the Word-Overlap Bias in Natural Language InferenceSara Rajaee, Yadollah Yaghoobzadeh, Mohammad Taher PilehvarEMNLP 2022 · 被引用 8 次
- Improving the robustness of NLI models with minimax trainingMichalis Korakakis, Andreas VlachosACL 2023 · 被引用 4 次
- Flexible Generation of Natural Language DeductionsKaj Bostrom, Xinyu Zhao, Swarat Chaudhuri, Greg DurrettEMNLP 2021 · 被引用 1 次
- Resolving Lexical Bias in Model EditingHammad Rizwan, Domenic Rosati, Ga Wu, Hassan SajjadICML 2025
它引用的顶会 Paper1
相关 Paper
- End-to-End Bias Mitigation by Modelling Biases in CorporaRabeeh Karimi Mahabadi, Yonatan Belinkov, James HendersonACL 2020 · 被引用 136 次
- Towards Stable Natural Language Understanding via Information Entropy Guided DebiasingLi Du, Xiao Ding, Zhouhao Sun, Ting Liu 等ACL 2023 · 被引用 1 次
- Towards Debiasing NLU Models from Unknown BiasesPrasetya Ajie Utama, Nafise Sadat Moosavi, Iryna GurevychEMNLP 2020 · 被引用 3 次
- Towards Understanding Gender Bias in Relation ExtractionAndrew Gaut, Tony Sun, Shirlyn Tang, Yuxin Huang 等ACL 2020 · 被引用 7 次
- A General Framework for Implicit and Explicit Debiasing of Distributional Word Vector SpacesAnne Lauscher, Goran Glavas, Simone Paolo Ponzetto, Ivan VulicAAAI 2020 · 被引用 68 次
