Reflect, Not Reflex: Inference-Based Common Ground Improves Dialogue Response Quality
Pei Zhou, Hyundong Cho, Pegah Jandaghi, Dong-Ho Lee, Bill Yuchen Lin, Jay Pujara, Xiang Ren
Abstract
Human communication relies on common ground (CG), the mutual knowledge and beliefs shared by participants, to produce coherent and interesting conversations. In this paper, we demonstrate that current response generation (RG) models produce generic and dull responses in dialogues because they act reflexively, failing to explicitly model CG, both due to the lack of CG in training data and the standard RG training procedure. We introduce Reflect, a dataset that annotates dialogues with explicit CG (materialized as inferences approximating shared knowledge and beliefs) and solicits 9k diverse human-generated responses each following one common ground. Using Reflect, we showcase the limitations of current dialogue data and RG models: less than half of the responses in current data is rated as high quality (sensible, specific, and interesting) and models trained using this data have even lower quality, while most Reflect responses are judged high quality. Next, we analyze whether CG can help models produce better quality responses by using Reflect CG to guide RG models. Surprisingly, we find that simply prompting GPT3 to "think" about CG generates 30% more quality responses, showing promising benefits to integrating CG into the RG process. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- SODA: Million-scale Dialogue Distillation with Social Commonsense ContextualizationHyunwoo Kim, Jack Hessel, Liwei Jiang, Peter West et al.EMNLP 2023 · 60 citations
- Navigating Rifts in Human-LLM Grounding: Study and BenchmarkOmar Shaikh, Hussein Mozannar, Gagan Bansal, Adam Fourney et al.ACL 2025 · 21 citations
- Can Large Language Models Be Good Companions?: An LLM-Based Eyewear System with Conversational Common GroundZhenyu Xu, Hailin Xu, Zhouyang Lu, Yingying Zhao et al.UbiComp 2024 · 20 citations
- I Cast Detect Thoughts: Learning to Converse and Guide with Intents and Theory-of-Mind in Dungeons and DragonsPei Zhou, Andrew Zhu, Jennifer Hu, Jay Pujara et al.ACL 2023 · 8 citations
- Grounded in Reality: Learning and Deploying Proactive LLM from Offline LogsFei Wei, Daoyuan Chen, Ce Wang, Yilun Huang et al.ICML 2026 · 2 citations
Builds on6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- (Comet-) Atomic 2020: On Symbolic and Neural Commonsense Knowledge GraphsJena D. Hwang, Chandra Bhagavatula, Ronan Le Bras, Jeff Da et al.AAAI 2021 · 458 citations
- MuTual: A Dataset for Multi-Turn Dialogue ReasoningLeyang Cui, Yu Wu, Shujie Liu, Yue Zhang et al.ACL 2020 · 115 citations
- Grounding Conversations with Improvised DialoguesHyundong Cho, Jonathan MayACL 2020 · 27 citations
Related papers
- An Annotated Corpus of Reference Resolution for Interpreting Common GroundingTakuma Udagawa, Akiko AizawaAAAI 2020 · 10 citations
- Call for Customized Conversation: Customized Conversation Grounding Persona and KnowledgeYoonna Jang, Jungwoo Lim, Yuna Hur, Dongsuk Oh et al.AAAI 2022 · 47 citations
- A Synthetic Data Generation Framework for Grounded DialoguesJianzhu Bao, Rui Wang, Yasheng Wang, Aixin Sun et al.ACL 2023 · 11 citations
- KPT: Keyword-Guided Pre-training for Grounded Dialog GenerationQi Zhu, Fei Mi, Zheng Zhang, Yasheng Wang et al.AAAI 2023 · 5 citations
- Think Before You Speak: Explicitly Generating Implicit Commonsense Knowledge for Response GenerationPei Zhou, Karthik Gopalakrishnan, Behnam Hedayatnia, Seokhwan Kim et al.ACL 2022 · 45 citations
