An Equal-Size Hard EM Algorithm for Diverse Dialogue Generation
Yuqiao Wen, Yongchang Hao, Yanshuai Cao, Lili Mou
摘要
Open-domain dialogue systems aim to interact with humans through natural language texts in an open-ended fashion. Despite the recent success of super large dialogue systems such as ChatGPT, using medium-to-small-sized dialogue systems remains the common practice as they are more lightweight and accessible; however, generating diverse dialogue responses is challenging, especially with smaller models. In this work, we propose an Equal-size Hard Expectation-Maximization (EqHard-EM) algorithm to train a multi-decoder model for diverse dialogue generation. Our algorithm assigns a sample to a decoder in a hard manner and additionally imposes an equal-assignment constraint to ensure that all decoders are well-trained. We provide detailed theoretical analysis to justify our approach. Further, experiments on two large-scale open-domain dialogue datasets verify that our EqHard-EM algorithm generates high-quality diverse responses 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Teacher Forcing Recovers Reward Functions for Text GenerationYongchang Hao, Yuxin Liu, Lili MouNeurIPS 2022 · 被引用 24 次
- f-Divergence Minimization for Sequence-Level Knowledge DistillationYuqiao Wen, Zichao Li, Wenyu Du, Lili MouACL 2023 · 被引用 14 次
- Ensemble Distillation for Unsupervised Constituency ParsingBehzad Shayegh, Yanshuai Cao, Xiaodan Zhu, Jackie C. K. Cheung 等ICLR 2024 · 被引用 10 次
- DiVERT: Distractor Generation with Variational Errors Represented as Text for Math Multiple-choice QuestionsNigel Fernandez, Alexander Scarlatos, Wanyong Feng, Simon Woodhead 等EMNLP 2024 · 被引用 10 次
- A Branching Decoder for Set GenerationZixian Huang, Gengyang Xiao, Yu Gu, Gong ChengICLR 2024 · 被引用 2 次
它引用的顶会 Paper13
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes 等ICLR 2020 · 被引用 4,112 次
- Expectation-Maximization Attention Networks for Semantic SegmentationXia Li, Zhisheng Zhong, Jianlong Wu, Yibo Yang 等ICCV 2019 · 被引用 639 次
- PLATO: Pre-trained Dialogue Generation Model with Discrete Latent VariableSiqi Bao, Huang He, Fan Wang, Hua Wu 等ACL 2020 · 被引用 229 次
- Optimus: Organizing Sentences via Pre-trained Modeling of a Latent SpaceChunyuan Li, Xiang Gao, Yuan Li, Baolin Peng 等EMNLP 2020 · 被引用 132 次
相关 Paper
- DialogVED: A Pre-trained Latent Variable Encoder-Decoder Model for Dialog Response GenerationWei Chen, Yeyun Gong, Song Wang, Bolun Yao 等ACL 2022
- EM Pre-training for Multi-party Dialogue Response GenerationYiyang Li, Hai ZhaoACL 2023 · 被引用 9 次
- Learning to Customize Model Structures for Few-shot Dialogue Generation TasksYiping Song, Zequn Liu, Wei Bi, Rui Yan 等ACL 2020 · 被引用 33 次
- Generating Dialogue Responses from a Semantic Latent SpaceWei-Jen Ko, Avik Ray, Yilin Shen, Hongxia JinEMNLP 2020 · 被引用 4 次
- Diversifying Dialog Generation via Adaptive Label SmoothingYida Wang, Yinhe Zheng, Yong Jiang, Minlie HuangACL 2021
