Generating Dialogue Responses from a Semantic Latent Space
Wei-Jen Ko, Avik Ray, Yilin Shen, Hongxia Jin
Abstract
Existing open-domain dialogue generation models are usually trained to mimic the gold response in the training set using cross-entropy loss on the vocabulary. However, a good response does not need to resemble the gold response, since there are multiple possible responses to a given prompt. In this work, we hypothesize that the current models are unable to integrate information from multiple semantically similar valid responses of a prompt, resulting in the generation of generic and uninformative responses. To address this issue, we propose an alternative to the end-to-end classification on vocabulary. We learn the pair relationship between the prompts and responses as a regression task on a latent space instead. In our novel dialog generation model, the representations of semantically related sentences are close to each other on the latent space. Human evaluation showed that learning the task on a continuous space can generate responses that are both relevant and informative.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c1e20099-5c0f-48e7-83bc-dccf23c74b4bBuilds on1
Related papers
- Evaluating Open-Domain Dialogues in Latent Space with Next Sentence Prediction and Mutual InformationKun Zhao, Bohao Yang, Chenghua Lin, Wenge Rong et al.ACL 2023 · 14 citations
- DialogVED: A Pre-trained Latent Variable Encoder-Decoder Model for Dialog Response GenerationWei Chen, Yeyun Gong, Song Wang, Bolun Yao et al.ACL 2022
- CoLV: A Collaborative Latent Variable Model for Knowledge-Grounded Dialogue GenerationHaolan Zhan, Lei Shen, Hongshen Chen, Hainan ZhangEMNLP 2021 · 15 citations
- Towards Efficient Dialogue Pre-training with Transferable and Interpretable Latent StructureXueliang Zhao, Lemao Liu, Tingchen Fu, Shuming Shi et al.EMNLP 2022 · 3 citations
- Seen to Unseen: Exploring Compositional Generalization of Multi-Attribute Controllable Dialogue GenerationWeihao Zeng, Lulu Zhao, Keqing He, Ruotong Geng et al.ACL 2023 · 1 citation
