Reinforcement Learning from Reformulations in Conversational Question Answering over Knowledge Graphs
Magdalena Kaiser, Rishiraj Saha Roy, Gerhard Weikum
Abstract
The rise of personal assistants has made conversational question answering (ConvQA) a very popular mechanism for user-system interaction. State-of-the-art methods for ConvQA over knowledge graphs (KGs) can only learn from crisp question-answer pairs found in popular benchmarks. In reality, however, such training data is hard to come by: users would rarely mark answers explicitly as correct or wrong. In this work, we take a step towards a more natural learning paradigm -from noisy and implicit feedback via question reformulations. A reformulation is likely to be triggered by an incorrect system response, whereas a new follow-up question could be a positive signal on the previous turn's answer. We present a reinforcement learning model, termed Conqer, that can learn from a conversational stream of questions and reformulations. Conqer models the answering process as multiple agents walking in parallel on the KG, where the walks are determined by actions sampled using a policy network. This policy network takes the question along with the conversational context as inputs and is trained via noisy rewards obtained from the reformulation likelihood. To evaluate Conqer, we create and release ConvRef, a benchmark with about 11𝑘 natural conversations containing around 205𝑘 reformulations. Experiments show that Conqer successfully learns to answer conversational questions from noisy reward signals, significantly improving over a state-of-the-art baseline.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a9910751-43ce-4243-bd45-dfb71e8fdd4fCited by top-tier papers10
- MMKGR: Multi-hop Multi-modal Knowledge Graph ReasoningShangfei Zheng, Weiqing Wang, Jianfeng Qu, Hongzhi Yin et al.ICDE 2023 · 40 citations
- Conversational Question Answering on Heterogeneous SourcesPhilipp Christmann, Rishiraj Saha Roy, Gerhard WeikumSIGIR 2022 · 26 citations
- Explainable Conversational Question Answering over Heterogeneous Sources via Iterative Graph Neural NetworksPhilipp Christmann, Rishiraj Saha Roy, Gerhard WeikumSIGIR 2023 · 21 citations
- Mixed Geometry Message and Trainable Convolutional Attention Network for Knowledge Graph CompletionBin Shang, Yinliang Zhao, Jun Liu, Di WangAAAI 2024 · 18 citations
- Double-Branch Multi-Attention based Graph Neural Network for Knowledge Graph CompletionHongcai Xu, Junpeng Bao, Wenbo LiuACL 2023 · 17 citations
Builds on3
- Open-Retrieval Conversational Question AnsweringChen Qu, Liu Yang, Cen Chen, Minghui Qiu et al.SIGIR 2020 · 84 citations
- Reinforced History Backtracking for Conversational Question AnsweringMinghui Qiu, Xinjing Huang, Cen Chen, Feng Ji et al.AAAI 2021 · 31 citations
- Message Passing for Hyper-Relational Knowledge GraphsMikhail Galkin, Priyansh Trivedi, Gaurav Maheshwari, Ricardo Usbeck et al.EMNLP 2020 · 17 citations
Related papers
- CONQRR: Conversational Query Rewriting for Retrieval with Reinforcement LearningZeqiu Wu, Yi Luan, Hannah Rashkin, David Reitter et al.EMNLP 2022 · 37 citations
- KBQA-R1: Reinforcing Large Language Models for Knowledge Base Question AnsweringXin Sun, Zhongqi Chen, Xing Zheng, Bowen Song et al.ICML 2026 · 2 citations
- Ditch the Gold Standard: Re-evaluating Conversational Question AnsweringHuihan Li, Tianyu Gao, Manan Goenka, Danqi ChenACL 2022 · 23 citations
- ChatR1: Reinforcement Learning for Conversational Reasoning and Retrieval Augmented Question AnsweringSimon Lupart, Mohammad Aliannejadi, Evangelos KanoulasACL 2026 · 5 citations
- CRFR: Improving Conversational Recommender Systems via Flexible Fragments Reasoning on Knowledge GraphsJinfeng Zhou, Bo Wang, Ruifang He, Yuexian HouEMNLP 2021 · 42 citations
