Dungeons and Dragons as a Dialog Challenge for Artificial Intelligence
Chris Callison-Burch, Gaurav Singh Tomar, Lara J. Martin, Daphne Ippolito, Suma Bailis, David Reitter
摘要
AI researchers have posited Dungeons and Dragons (D&D) as a challenge problem to test systems on various language-related capabilities. In this paper, we frame D&D specifically as a dialogue system challenge, where the tasks are to both generate the next conversational turn in the game and predict the state of the game given the dialogue history. We create a gameplay dataset consisting of nearly 900 games, with a total of 7,000 players, 800,000 dialogue turns, 500,000 dice rolls, and 58 million words. We automatically annotate the data with partial state information about the game play. We train a large language model (LM) to generate the next game turn, conditioning it on different information. The LM can respond as a particular character or as the player who runs the game—i.e., the Dungeon Master (DM). It is trained to produce dialogue that is either in-character (roleplaying in the fictional world) or out-of-character (discussing rules or strategy). We perform a human evaluation to determine what factors make the generated output plausible and interesting. We further perform an automatic evaluation to determine how well the model can predict the game state given the history and examine how well tracking the game state improves its ability to produce plausible conversational output.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris 等UIST 2023 · 被引用 1,882 次
- Character-LLM: A Trainable Agent for Role-PlayingYunfan Shao, Linyang Li, Junqi Dai, Xipeng QiuEMNLP 2023 · 被引用 97 次
- I Cast Detect Thoughts: Learning to Converse and Guide with Intents and Theory-of-Mind in Dungeons and DragonsPei Zhou, Andrew Zhu, Jennifer Hu, Jay Pujara 等ACL 2023 · 被引用 8 次
- SimSpark: Interactive Simulation of Social Media BehaviorsZiyue Lin, Yi Shan, Lin Gao, Xinghua Jia 等CSCW 2025 · 被引用 5 次
- Ontologically Faithful Generation of Non-Player Character DialoguesNathaniel Weir, Ryan Thomas, Randolph D'Amore, Kellie Hill 等EMNLP 2024 · 被引用 2 次
它引用的顶会 Paper7
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- PlotMachines: Outline-Conditioned Generation with Dynamic Plot State TrackingHannah Rashkin, Asli Celikyilmaz, Yejin Choi, Jianfeng GaoEMNLP 2020 · 被引用 100 次
- Keep CALM and Explore: Language Models for Action Generation in Text-based GamesShunyu Yao, Rohan Rao, Matthew J. Hausknecht, Karthik NarasimhanEMNLP 2020 · 被引用 67 次
- Hierarchical Reinforcement Learning for Open-Domain DialogAbdelrhman Saleh, Natasha Jaques, Asma Ghandeharioun, Judy Hanwen Shen 等AAAI 2020 · 被引用 60 次
- Storytelling with Dialogue: A Critical Role Dungeons and Dragons DatasetRevanth Rameshkumar, Peter BaileyACL 2020 · 被引用 37 次
相关 Paper
- FIREBALL: A Dataset of Dungeons and Dragons Actual-Play with Structured Game State InformationAndrew Zhu, Karmanya Aggarwal, Alexander H. Feng, Lara J. Martin 等ACL 2023 · 被引用 9 次
- Personalized Quest and Dialogue Generation in Role-Playing Games: A Knowledge Graph- and Language Model-based ApproachTrevor Ashby, Braden K. Webb, Gregory Knapp, Jackson Searle 等CHI 2023 · 被引用 64 次
- DMT-RoleBench: A Dynamic Multi-Turn Dialogue Based Benchmark for Role-Playing Evaluation of Large Language Model and AgentDingbo Yuan, Yipeng Chen, Guodong Liu, Chenchen Li 等AAAI 2025 · 被引用 6 次
- Don't Stop the Multi-Party! On Generating Synthetic Written Multi-Party Conversations with ConstraintsNicolò Penzo, Marco Guerini, Bruno Lepri, Goran Glavas 等AAAI 2026 · 被引用 3 次
- LMRL Gym: Benchmarks for Multi-Turn Reinforcement Learning with Language ModelsMarwa Abdulhai, Isadora White, Charlie Victor Snell, Charles Sun 等ICML 2025
