Data Manipulation: Towards Effective Instance Learning for Neural Dialogue Generation via Learning to Augment and Reweight
Hengyi Cai, Hongshen Chen, Yonghao Song, Cheng Zhang, Xiaofang Zhao, Dawei Yin
Abstract
Current state-of-the-art neural dialogue models learn from human conversations following the data-driven paradigm. As such, a reliable training corpus is the crux of building a robust and well-behaved dialogue model. However, due to the open-ended nature of human conversations, the quality of user-generated training data varies greatly, and effective training samples are typically insufficient while noisy samples frequently appear. This impedes the learning of those data-driven neural dialogue models. Therefore, effective dialogue learning requires not only more reliable learning samples, but also fewer noisy samples. In this paper, we propose a data manipulation framework to proactively reshape the data distribution towards reliable samples by augmenting and highlighting effective learning samples as well as reducing the effect of inefficient samples simultaneously. In particular, the data manipulation model selectively augments the training samples and assigns an importance weight to each instance to reform the training data. Note that, the proposed data manipulation framework is fully data-driven and learnable. It not only manipulates training samples to optimize the dialogue generation model, but also learns to increase its manipulation skills through gradient descent with validation samples. Extensive experiments show that our framework can improve the dialogue generation performance with respect to various automatic evaluation metrics and human judgments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Physics of Language Models: Part 3.1, Knowledge Storage and ExtractionZeyuan Allen-Zhu, Yuanzhi LiICML 2024 · 258 citations
- Dialogue Distillation: Open-Domain Dialogue Augmentation Using Unpaired DataRongsheng Zhang, Yinhe Zheng, Jianzhi Shao, Xiaoxi Mao et al.EMNLP 2020 · 25 citations
- Text AutoAugment: Learning Compositional Augmentation Policy for Text ClassificationShuhuai Ren, Jinchao Zhang, Lei Li, Xu Sun et al.EMNLP 2021 · 22 citations
- Counterfactual Data Augmentation via Perspective Transition for Open-Domain DialoguesJiao Ou, Jinchao Zhang, Yang Feng, Jie ZhouEMNLP 2022 · 9 citations
- Learning towards Selective Data Augmentation for Dialogue GenerationXiuying Chen, Mingzhe Li, Jiayi Zhang, Xiaoqiang Xia et al.AAAI 2023 · 7 citations
Builds on2
Related papers
- A Model-agnostic Data Manipulation Method for Persona-based Dialogue GenerationYu Cao, Wei Bi, Meng Fang, Shuming Shi et al.ACL 2022
- Paraphrase Augmented Task-Oriented Dialog GenerationSilin Gao, Yichi Zhang, Zhijian Ou, Zhou YuACL 2020 · 78 citations
- Learning from Noisy Labels via Self-Taught On-the-Fly Meta Loss RescalingMichael Heck, Christian Geishauser, Nurul Lubis, Carel van Niekerk et al.AAAI 2025 · 3 citations
- Negative Training for Neural Dialogue Response GenerationTianxing He, James R. GlassACL 2020 · 52 citations
- Dialog State Tracking with Reinforced Data AugmentationYichun Yin, Lifeng Shang, Xin Jiang, Xiao Chen et al.AAAI 2020 · 25 citations
