Logical Reasoning with Span-Level Predictions for Interpretable and Robust NLI Models
Joe Stacey, Pasquale Minervini, Haim Dubossarsky, Marek Rei
摘要
Current Natural Language Inference (NLI) models achieve impressive results, sometimes outperforming humans when evaluating on in-distribution test sets. However, as these models are known to learn from annotation artefacts and dataset biases, it is unclear to what extent the models are learning the task of NLI instead of learning from shallow heuristics in their training data.We address this issue by introducing a logical reasoning framework for NLI, creating highly transparent model decisions that are based on logical rules. Unlike prior work, we show that improved interpretability can be achieved without decreasing the predictive accuracy. We almost fully retain performance on SNLI, while also identifying the exact hypothesis spans that are responsible for each model prediction.Using the e-SNLI human explanations, we verify that our model makes sensible decisions at a span level, despite not using any span labels during training. We can further improve model performance and the span-level decisions by using the e-SNLI explanations during training. Finally, our model is more robust in a reduced data setting. When training with only 1,000 examples, out-of-distribution performance improves on the MNLI matched and mismatched validation sets by 13% and 16% relative to the baseline. Training with fewer observations yields further improvements, both in-distribution and out-of-distribution.<br/>
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- QA-NatVer: Question Answering for Natural Logic-based Fact VerificationRami Aly, Marek Strong, Andreas VlachosEMNLP 2023 · 被引用 7 次
- Extractive Fact Decomposition for Interpretable Natural Language Inference in one Forward PassNicholas Popovic, Michael FärberEMNLP 2025 · 被引用 1 次
- InfoLossQA: Characterizing and Recovering Information Loss in Text SimplificationJan Trienes, Sebastian Joseph, Jörg Schlötterer, Christin Seifert 等ACL 2024
- Language Models, Graph Searching, and Supervision Adulteration: When More Supervision is Less and How to Make More MoreArvid FrydenlundACL 2025
它引用的顶会 Paper9
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- Adversarial NLI: A New Benchmark for Natural Language UnderstandingYixin Nie, Adina Williams, Emily Dinan, Mohit Bansal 等ACL 2020 · 被引用 602 次
- End-to-End Bias Mitigation by Modelling Biases in CorporaRabeeh Karimi Mahabadi, Yonatan Belinkov, James HendersonACL 2020 · 被引用 136 次
- Leap-Of-Thought: Teaching Pre-Trained Models to Systematically Reason Over Implicit KnowledgeAlon Talmor, Oyvind Tafjord, Peter Clark, Yoav Goldberg 等NeurIPS 2020 · 被引用 119 次
- Supervising Model Attention with Human Explanations for Robust Natural Language InferenceJoe Stacey, Yonatan Belinkov, Marek ReiAAAI 2022 · 被引用 52 次
相关 Paper
- LIREx: Augmenting Language Inference with Relevant ExplanationsXinyan Zhao, V. G. Vinod VydiswaranAAAI 2021 · 被引用 41 次
- Weakly Supervised Explainable Phrasal Reasoning with Neural Fuzzy LogicZijun Wu, Zi Xuan Zhang, Atharva Naik, Zhijian Mei 等ICLR 2023 · 被引用 5 次
- Rule Discovery for Natural Language Inference Data Generation Using Out-of-Distribution DetectionJuyoung Han, Hyunsun Hwang, Changki LeeEMNLP 2025
- Improving the robustness of NLI models with minimax trainingMichalis Korakakis, Andreas VlachosACL 2023 · 被引用 4 次
- FLamE: Few-shot Learning from Natural Language ExplanationsYangqiaoyu Zhou, Yiming Zhang, Chenhao TanACL 2023 · 被引用 8 次
