Model-Based Simulation for Optimising Smart Reply
Benjamin Towle, Ke Zhou
摘要
Smart Reply (SR) systems present a user with a set of replies, of which one can be selected in place of having to type out a response. To perform well at this task, a system should be able to effectively present the user with a diverse set of options, to maximise the chance that at least one of them conveys the user’s desired response. This is a significant challenge, due to the lack of datasets containing sets of responses to learn from. Resultantly, previous work has focused largely on post-hoc diversification, rather than explicitly learning to predict sets of responses. Motivated by this problem, we present a novel method SimSR, that employs model-based simulation to discover high-value response sets, through simulating possible user responses with a learned world model. Unlike previous approaches, this allows our method to directly optimise the end-goal of SR–maximising the relevance of at least one of the predicted replies. Empirically on two public datasets, when compared to SoTA baselines, our method achieves up to 21% and 18% improvement in ROUGE score and Self-ROUGE score respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Poly-encoders: Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence ScoringSamuel Humeau, Kurt Shuster, Marie-Anne Lachaux, Jason WestonICLR 2020 · 被引用 316 次
- Monte-Carlo Planning and Learning with Language Action Value EstimatesYoungsoo Jang, Seokin Seo, Jongmin Lee, Kee-Eung KimICLR 2021 · 被引用 18 次
- A Dataset and Baselines for Multilingual Reply SuggestionMozhi Zhang, Wei Wang, Budhaditya Deb, Guoqing Zheng 等ACL 2021
相关 Paper
- AvgOut: A Simple Output-Probability Measure to Eliminate Dull ResponsesTong Niu, Mohit BansalAAAI 2020 · 被引用 3 次
- Evaluating Conversational Recommender Systems via User SimulationShuo Zhang, Krisztian BalogKDD 2020 · 被引用 80 次
- Learning Neural Templates for Recommender Dialogue SystemZujie Liang, Huang Hu, Can Xu, Jian Miao 等EMNLP 2021 · 被引用 40 次
- Exploiting Simulated User Feedback for Conversational Search: Ranking, Rewriting, and BeyondPaul Owoicho, Ivan Sekulic, Mohammad Aliannejadi, Jeffrey Dalton 等SIGIR 2023 · 被引用 31 次
- A Textual Dataset for Situated Proactive Response SelectionNaoki Otani, Jun Araki, HyeongSik Kim, Eduard H. HovyACL 2023 · 被引用 1 次
