RetGen: A Joint Framework for Retrieval and Grounded Text Generation Modeling
Yizhe Zhang, Siqi Sun, Xiang Gao, Yuwei Fang, Chris Brockett, Michel Galley, Jianfeng Gao, Bill Dolan
Abstract
Recent advances in large-scale pre-training such as GPT-3 allow seemingly high quality text to be generated from a given prompt. However, such generation systems often suffer from problems of hallucinated facts, and are not inherently designed to incorporate useful external information. Grounded generation models appear to offer remedies, but their training typically relies on rarely-available parallel data where information-relevant documents are provided for context. We propose a framework that alleviates this data constraint by jointly training a grounded generator and document retriever on the language model signal. The model learns to reward retrieval of the documents with the highest utility in generation, and attentively combines them using a Mixture-of-Experts (MoE) ensemble to generate follow-on text. We demonstrate that both generator and retriever can take advantage of this joint training and work synergistically to produce more informative and relevant text in both prose and dialogue generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1167a18d-65d1-4e4a-b507-a3b9cec53ab8Cited by top-tier papers3
- RepoCoder: Repository-Level Code Completion Through Iterative Retrieval and GenerationFengji Zhang, Bei Chen, Yue Zhang, Jacky Keung et al.EMNLP 2023 · 110 citations
- KPT: Keyword-Guided Pre-training for Grounded Dialog GenerationQi Zhu, Fei Mi, Zheng Zhang, Yasheng Wang et al.AAAI 2023 · 5 citations
- Diversify Question Generation with Retrieval-Augmented Style TransferQi Gou, Zehua Xia, Bowen Yu, Haiyang Yu et al.EMNLP 2023 · 4 citations
Builds on6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text RetrievalLee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang et al.ICLR 2021 · 1,547 citations
- Pre-training via ParaphrasingMike Lewis, Marjan Ghazvininejad, Gargi Ghosh, Armen Aghajanyan et al.NeurIPS 2020 · 165 citations
- Knowledge-Grounded Dialogue Generation with Pre-trained Language ModelsXueliang Zhao, Wei Wu, Can Xu, Chongyang Tao et al.EMNLP 2020 · 153 citations
- Knowledge Graph-Augmented Abstractive Summarization with Semantic-Driven Cloze RewardLuyang Huang, Lingfei Wu, Lu WangACL 2020 · 152 citations
Related papers
- Hindsight: Posterior-guided training of retrievers for improved open-ended generationAshwin Paranjape, Omar Khattab, Christopher Potts, Matei Zaharia et al.ICLR 2022 · 48 citations
- A Pre-training Strategy for Zero-Resource Response Selection in Knowledge-Grounded ConversationsChongyang Tao, Changyu Chen, Jiazhan Feng, Ji-Rong Wen et al.ACL 2021
- A Synthetic Data Generation Framework for Grounded DialoguesJianzhu Bao, Rui Wang, Yasheng Wang, Aixin Sun et al.ACL 2023 · 11 citations
- Eliciting Knowledge from Large Pre-Trained Models for Unsupervised Knowledge-Grounded ConversationYanyang Li, Jianqiao Zhao, Michael R. Lyu, Liwei WangEMNLP 2022 · 11 citations
- Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog GenerationJinyoung Park, Minseok Joo, Joo-Kyung Kim, Hyunwoo J. KimEMNLP 2024 · 3 citations
