Robustifying Multi-hop QA through Pseudo-Evidentiality Training
Kyungjae Lee, Seung-won Hwang, Sang-eun Han, Dohyeon Lee
Abstract
This paper studies the bias problem of multihop question answering models, of answering correctly without correct reasoning. One way to robustify these models is by supervising to not only answer right, but also with right reasoning chains. An existing direction is to annotate reasoning chains to train models, requiring expensive additional annotations. In contrast, we propose a new approach to learn evidentiality, deciding whether the answer prediction is supported by correct evidences, without such annotations. Instead, we compare counterfactual changes in answer confidence with and without evidence sentences, to generate "pseudo-evidentiality" annotations. We validate our proposed model on an original set and challenge set in HotpotQA, showing that our method is accurate and robust in multi-hop reasoning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 929287ea-c293-4603-8734-8f2409f56cd5Cited by top-tier papers6
- Merging Generated and Retrieved Knowledge for Open-Domain QAYunxiang Zhang, Muhammad Khalifa, Lajanugen Logeswaran, Moontae Lee et al.EMNLP 2023 · 12 citations
- Why LLMs Hallucinate, and How to Get (Evidential) Closure: Perceptual, Intensional, and Extensional Learning for Faithful Natural Language GenerationAdam BouyamournEMNLP 2023 · 9 citations
- Teaching Broad Reasoning Skills for Multi-Step QA by Generating Hard ContextsHarsh Trivedi, Niranjan Balasubramanian, Tushar Khot, Ashish SabharwalEMNLP 2022 · 6 citations
- Counterfactual Multihop QA: A Cause-Effect Approach for Reducing Disconnected ReasoningWangzhen Guo, Qinkang Gong, Yanghui Rao, Hanjiang LaiACL 2023 · 4 citations
- Single Sequence Prediction over Reasoning Graphs for Multi-hop QAGowtham Ramesh, Makesh Narsimhan Sreedhar, Junjie HuACL 2023 · 1 citation
Builds on7
- Learning to Retrieve Reasoning Paths over Wikipedia Graph for Question AnsweringAkari Asai, Kazuma Hashimoto, Hannaneh Hajishirzi, Richard Socher et al.ICLR 2020 · 322 citations
- Hierarchical Graph Network for Multi-hop Question AnsweringYuwei Fang, Siqi Sun, Zhe Gan, Rohit Pillai et al.EMNLP 2020 · 157 citations
- A Self-Training Method for Machine Reading Comprehension with Soft Evidence ExtractionYilin Niu, Fangkai Jiao, Mantong Zhou, Ting Yao et al.ACL 2020 · 33 citations
- Mind the Trade-off: Debiasing NLU Models without Degrading the In-distribution PerformancePrasetya Ajie Utama, Nafise Sadat Moosavi, Iryna GurevychACL 2020 · 11 citations
- Is Multihop QA in DiRe Condition? Measuring and Reducing Disconnected ReasoningHarsh Trivedi, Niranjan Balasubramanian, Tushar Khot, Ashish SabharwalEMNLP 2020 · 3 citations
Related papers
- Hop, Union, Generate: Explainable Multi-hop Reasoning without Rationale SupervisionWenting Zhao, Justin T. Chiu, Claire Cardie, Alexander M. RushEMNLP 2023 · 4 citations
- Generating Multi-hop Reasoning Questions to Improve Machine Reading ComprehensionJianxing Yu, Xiaojun Quan, Qinliang Su, Jian YinWWW 2020 · 25 citations
- F1 is Not Enough! Models and Evaluation Towards User-Centered Explainable Question AnsweringHendrik Schuff, Heike Adel, Ngoc Thang VuEMNLP 2020
- Low-Resource Generation of Multi-hop Reasoning QuestionsJianxing Yu, Wei Liu, Shuang Qiu, Qinliang Su et al.ACL 2020 · 11 citations
- DeFacto: Counterfactual Thinking with Images for Enforcing Evidence-Grounded and Faithful ReasoningTianrun Xu, Haoda Jing, Ye Li, Yuquan Wei et al.ICML 2026 · 8 citations
