Learning from Perturbations: Diverse and Informative Dialogue Generation with Inverse Adversarial Training
Wangchunshu Zhou, Qifei Li, Chenle Li
摘要
In this paper, we propose Inverse Adversarial Training (IAT) algorithm for training neural dialogue systems to avoid generic responses and model dialogue history better. In contrast to standard adversarial training algorithms, IAT encourages the model to be sensitive to the perturbation in the dialogue history and therefore learning from perturbations. By giving higher rewards for responses whose output probability reduces more significantly when dialogue history is perturbed, the model is encouraged to generate more diverse and consistent responses. By penalizing the model when generating the same response given perturbed dialogue history, the model is forced to better capture dialogue history and generate more informative responses. Experimental results on two benchmark datasets show that our approach can better model dialogue history and generate more diverse and consistent responses. In addition, we point out a problem of the widely used maximum mutual information (MMI) based methods for improving the diversity of dialogue response generation models and demonstrate it empirically.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Factual and Informative Review Generation for Explainable RecommendationZhouhang Xie, Sameer Singh, Julian J. McAuley, Bodhisattwa Prasad MajumderAAAI 2023 · 被引用 36 次
- Evade the Trap of Mediocrity: Promoting Diversity and Novelty in Text Generation via Concentrating AttentionWenhao Li, Xiaoyuan Yi, Jinyi Hu, Maosong Sun 等EMNLP 2022 · 被引用 1 次
它引用的顶会 Paper3
- Diverse and Informative Dialogue Generation with Context-Specific Commonsense Knowledge AwarenessSixing Wu, Ying Li, Dawei Zhang, Yang Zhou 等ACL 2020 · 被引用 104 次
- Learning to Compare for Better Training and Evaluation of Open Domain Natural Language Generation ModelsWangchunshu Zhou, Ke XuAAAI 2020 · 被引用 49 次
- Self-Adversarial Learning with Comparative Discrimination for Text GenerationWangchunshu Zhou, Tao Ge, Ke Xu, Furu Wei 等ICLR 2020 · 被引用 20 次
相关 Paper
- Learning towards Selective Data Augmentation for Dialogue GenerationXiuying Chen, Mingzhe Li, Jiayi Zhang, Xiaoqiang Xia 等AAAI 2023 · 被引用 7 次
- Curriculum Prompt Learning with Self-Training for Abstractive Dialogue SummarizationChangqun Li, Linlin Wang, Xin Lin, Gerard de Melo 等EMNLP 2022 · 被引用 8 次
- Negative Training for Neural Dialogue Response GenerationTianxing He, James R. GlassACL 2020 · 被引用 52 次
- ProphetChat: Enhancing Dialogue Generation with Simulation of Future ConversationChang Liu, Xu Tan, Chongyang Tao, Zhenxin Fu 等ACL 2022
- Counterfactual Off-Policy Training for Neural Dialogue GenerationQingfu Zhu, Wei-Nan Zhang, Ting Liu, William Yang WangEMNLP 2020 · 被引用 18 次
