Model-Based Simulation for Optimising Smart Reply
Benjamin Towle, Ke Zhou
Abstract
Smart Reply (SR) systems present a user with a set of replies, of which one can be selected in place of having to type out a response. To perform well at this task, a system should be able to effectively present the user with a diverse set of options, to maximise the chance that at least one of them conveys the user’s desired response. This is a significant challenge, due to the lack of datasets containing sets of responses to learn from. Resultantly, previous work has focused largely on post-hoc diversification, rather than explicitly learning to predict sets of responses. Motivated by this problem, we present a novel method SimSR, that employs model-based simulation to discover high-value response sets, through simulating possible user responses with a learned world model. Unlike previous approaches, this allows our method to directly optimise the end-goal of SR–maximising the relevance of at least one of the predicted replies. Empirically on two public datasets, when compared to SoTA baselines, our method achieves up to 21% and 18% improvement in ROUGE score and Self-ROUGE score respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f37984e2-8dda-49e3-9b1a-9f4ccedb6f8aBuilds on3
- Poly-encoders: Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence ScoringSamuel Humeau, Kurt Shuster, Marie-Anne Lachaux, Jason WestonICLR 2020 · 316 citations
- Monte-Carlo Planning and Learning with Language Action Value EstimatesYoungsoo Jang, Seokin Seo, Jongmin Lee, Kee-Eung KimICLR 2021 · 18 citations
- A Dataset and Baselines for Multilingual Reply SuggestionMozhi Zhang, Wei Wang, Budhaditya Deb, Guoqing Zheng et al.ACL 2021
Related papers
- AvgOut: A Simple Output-Probability Measure to Eliminate Dull ResponsesTong Niu, Mohit BansalAAAI 2020 · 3 citations
- Evaluating Conversational Recommender Systems via User SimulationShuo Zhang, Krisztian BalogKDD 2020 · 80 citations
- Learning Neural Templates for Recommender Dialogue SystemZujie Liang, Huang Hu, Can Xu, Jian Miao et al.EMNLP 2021 · 40 citations
- Exploiting Simulated User Feedback for Conversational Search: Ranking, Rewriting, and BeyondPaul Owoicho, Ivan Sekulic, Mohammad Aliannejadi, Jeffrey Dalton et al.SIGIR 2023 · 31 citations
- A Textual Dataset for Situated Proactive Response SelectionNaoki Otani, Jun Araki, HyeongSik Kim, Eduard H. HovyACL 2023 · 1 citation
