Can You Put it All Together: Evaluating Conversational Agents' Ability to Blend Skills
Eric Michael Smith, Mary Williamson, Kurt Shuster, Jason Weston, Y-Lan Boureau
摘要
Being engaging, knowledgeable, and empathetic are all desirable general qualities in a conversational agent. Previous work has introduced tasks and datasets that aim to help agents to learn those qualities in isolation and gauge how well they can express them. But rather than being specialized in one single quality, a good open-domain conversational agent should be able to seamlessly blend them all into one cohesive conversational flow. In this work, we investigate several ways to combine models trained towards isolated capabilities, ranging from simple model aggregation schemes that require minimal additional training, to various forms of multi-task training that encompass several skills at all training stages. We further propose a new dataset, Blended-SkillTalk, to analyze how these capabilities would mesh together in a natural conversation, and compare the performance of different architectures and training schemes. Our experiments show that multi-tasking over several tasks that focus on particular capabilities results in better blended conversation performance compared to models trained on a single skill, and that both unified or two-stage approaches perform well if they are constructed to avoid unwanted bias in skill selection or are fine-tuned on our new task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper48
- Fusing Task-Oriented and Open-Domain Dialogues in Conversational AgentsTom Young, Frank Xing, Vlad Pandelea, Jinjie Ni 等AAAI 2022 · 被引用 64 次
- SODA: Million-scale Dialogue Distillation with Social Commonsense ContextualizationHyunwoo Kim, Jack Hessel, Liwei Jiang, Peter West 等EMNLP 2023 · 被引用 60 次
- Perspective-taking and Pragmatics for Generating Empathetic Responses Focused on Emotion CausesHyunwoo Kim, Byeongchang Kim, Gunhee KimEMNLP 2021 · 被引用 57 次
- "I'm sorry to hear that": Finding New Biases in Language Models with a Holistic Descriptor DatasetEric Michael Smith, Melissa Hall, Melanie Kambadur, Eleonora Presani 等EMNLP 2022 · 被引用 56 次
- SaFeRDialogues: Taking Feedback Gracefully after Conversational Safety FailuresMegan Ung, Jing Xu, Y-Lan BoureauACL 2022 · 被引用 54 次
相关 Paper
- BotsTalk: Machine-sourced Framework for Automatic Curation of Large-scale Multi-skill Dialogue DatasetsMinju Kim, Chaehyeong Kim, Yongho Song, Seung-won Hwang 等EMNLP 2022 · 被引用 8 次
- The Dialogue Dodecathlon: Open-Domain Knowledge and Image Grounded Conversational AgentsKurt Shuster, Da Ju, Stephen Roller, Emily Dinan 等ACL 2020 · 被引用 9 次
- Multi-Source Probing for Open-Domain Conversational UnderstandingYuanxi Li, Hao Zhou, Jie Zhou, Minlie HuangEMNLP 2023
- Generative Expressive Conversational Speech SynthesisRui Liu, Yifan Hu, Yi Ren, Xiang Yin 等ACM MM 2024 · 被引用 15 次
- Beyond Task-Oriented and Chitchat Dialogues: Proactive and Transition-Aware Conversational AgentsYejin Yoon, Yuri Son, Namyoung So, Minseo Kim 等EMNLP 2025
