Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
Nathaniel Weir, Kate Sanders, Orion Weller, Shreya Sharma, Dongwei Jiang, Zhengping Jiang, Bhavana Dalvi Mishra, Oyvind Tafjord, Peter A. Jansen, Peter Clark, Benjamin Van Durme
摘要
Recent language models enable new opportunities for structured reasoning with text, such as the construction of intuitive, proof-like textual entailment trees without relying on brittle formal logic (Tafjord et al., 2022; Weir et al., 2024) . However, progress in this direction has been hampered by a long-standing lack of a clear protocol for determining what valid compositional entailment is. This absence causes noisy datasets and limited performance gains by modern neuro-symbolic engines. To address these problems, we formulate a consistent and theoretically grounded approach to annotating decompositional entailment and evaluate its impact on LLM-based textual inference. We find that our new dataset, RDTE (Recognizing Decompositional Textual Entailment), has a substantially higher internal consistency (+9%) than prior decompositional entailment datasets. We also find that training an RDTE-oriented entailment classifier via knowledge distillation and employing it in an entailment tree reasoning engine significantly improves both accuracy and proof quality, illustrating the practical benefit of this advance for textual inference.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language TextsHelia Hashemi, Jason Eisner, Corby Rosset, Benjamin Van Durme 等ACL 2024 · 被引用 27 次
- Probabilistic Soundness Guarantees in LLM Reasoning ChainsWeiqiu You, Anton Xue, Shreya Havaldar, Delip Rao 等EMNLP 2025 · 被引用 9 次
- Bonsai: Interpretable Tree-Adaptive Grounded ReasoningKate Sanders, Benjamin Van DurmeAAAI 2026 · 被引用 1 次
- Assessing Reliability and Political Bias In LLMs' Judgements of Formal and Material Inferences With Partisan ConclusionsReto Gubelmann, Ghassen KarrayACL 2025
- RATIONALYST: Pre-training Process-Supervision for Improving ReasoningDongwei Jiang, Guoxuan Wang, Yining Lu, Andrew Wang 等ACL 2025
它引用的顶会 Paper8
- QASC: A Dataset for Question Answering via Sentence CompositionTushar Khot, Peter Clark, Michal Guerquin, Peter Jansen 等AAAI 2020 · 被引用 387 次
- Entailer: Answering Questions with Faithful and Truthful Chains of ReasoningOyvind Tafjord, Bhavana Dalvi Mishra, Peter ClarkEMNLP 2022 · 被引用 28 次
- Generating Natural Language Proofs with Verifier-Guided SearchKaiyu Yang, Jia Deng, Danqi ChenEMNLP 2022 · 被引用 22 次
- Faithful Question Answering with Monte-Carlo PlanningRuixin Hong, Hongming Zhang, Hong Zhao, Dong Yu 等ACL 2023 · 被引用 7 次
- Explaining Answers with Entailment TreesBhavana Dalvi, Peter Jansen, Oyvind Tafjord, Zhengnan Xie 等EMNLP 2021 · 被引用 6 次
相关 Paper
- Structured Reasoning for LLMs: A Unified Framework for Efficiency and ExplainabilityYubo Dong, Hehe Fan, Linchao Zhu, Yi YangICLR 2026
- Logically Consistent Language Models via Neuro-Symbolic IntegrationDiego Calanzone, Stefano Teso, Antonio VergariICLR 2025 · 被引用 2 次
- Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge DistillationSiddhant Bhambri, Upasana Biswas, Subbarao KambhampatiACL 2026 · 被引用 9 次
- Disentangling Memory and Reasoning Ability in Large Language ModelsMingyu Jin, Weidi Luo, Sitao Cheng, Xinyi Wang 等ACL 2025
- RLET: A Reinforcement Learning Based Approach for Explainable QA with Entailment TreesTengxiao Liu, Qipeng Guo, Xiangkun Hu, Yue Zhang 等EMNLP 2022 · 被引用 8 次
