Weakly-Supervised Question Answering with Effective Rank and Weighted Loss over Candidates
Haozhe Qin, Jiangang Zhu, Beijun Shen
摘要
We study the weakly supervised question answering problem. Weak-ly supervised question answering aims to learn how the questions should be answered directly from the pairs without golden solutions/evidences, which makes question answering models much easier to scale to many domains. However, in weak supervision setup, a question typically involves many candidate solutions and the spuriousness of candidate solutions will hurt the performance of the question answering models. In this paper, we present an effective method to learn a question answering model in a weak supervision way. Specifically, in order to reduce the spuriousness of candidate solutions used for training, we design several simple yet effective scoring functions to rank the candidate solutions. Despite its simplicity, this ranking process can improve the quality of the training data significantly with fewer spurious candidates left. Then, different from previous approaches that either treat all candidates equally for training or only select the candidate with the largest likelihood in each iteration, we formulate this problem as a multi-task learning problem by weighing the losses computed from top-k candidates. Experimental results show that, our method1 can outperform previous approaches on both semantic parsing and machine reading comprehension tasks.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- A Mutual Information Maximization Approach for the Spurious Solution Problem in Weakly Supervised Question AnsweringZhihong Shao, Lifeng Shang, Qun Liu, Minlie HuangACL 2021
- Multi-Task Learning with Generative Adversarial Training for Multi-Passage Machine Reading ComprehensionQiyu Ren, Xiang Cheng, Sen SuAAAI 2020 · 被引用 15 次
- Improving Passage Retrieval with Zero-Shot Question GenerationDevendra Singh Sachan, Mike Lewis, Mandar Joshi, Armen Aghajanyan 等EMNLP 2022 · 被引用 69 次
- Multi-source Meta Transfer for Low Resource Multiple-Choice Question AnsweringMing Yan, Hao Zhang, Di Jin, Joey Tianyi ZhouACL 2020 · 被引用 21 次
- ReSCORE: Label-free Iterative Retriever Training for Multi-hop Question Answering with Relevance-Consistency SupervisionDosung Lee, Wonjun Oh, Boyoung Kim, Minyoung Kim 等ACL 2025
