Situated Dialogue Learning through Procedural Environment Generation
Prithviraj Ammanabrolu, Renee Jia, Mark O. Riedl
摘要
We teach goal-driven agents to interactively act and speak in situated environments by training on generated curriculums. Our agents operate in LIGHT (Urbanek et al., 2019)-a large-scale crowd-sourced fantasy text adventure game wherein an agent perceives and interacts with the world through textual natural language. Goals in this environment take the form of character-based quests, consisting of personas and motivations. We augment LIGHT by learning to procedurally generate additional novel textual worlds and quests to create a curriculum of steadily increasing difficulty for training agents to achieve such goals. In particular, we measure curriculum difficulty in terms of the rarity of the quest in the original training distribution-an easier environment is one that is more likely to have been found in the unaugmented dataset. An ablation study shows that this method of learning from the tail of a distribution results in significantly higher generalization abilities as measured by zeroshot performance on never-before-seen quests.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Neural Theory-of-Mind? On the Limits of Social Intelligence in Large LMsMaarten Sap, Ronan Le Bras, Daniel Fried, Yejin ChoiEMNLP 2022 · 被引用 92 次
- I Cast Detect Thoughts: Learning to Converse and Guide with Intents and Theory-of-Mind in Dungeons and DragonsPei Zhou, Andrew Zhu, Jennifer Hu, Jay Pujara 等ACL 2023 · 被引用 8 次
- Ontologically Faithful Generation of Non-Player Character DialoguesNathaniel Weir, Ryan Thomas, Randolph D'Amore, Kellie Hill 等EMNLP 2024 · 被引用 2 次
- ScienceWorld: Is your Agent Smarter than a 5th Grader?Ruoyao Wang, Peter A. Jansen, Marc-Alexandre Côté, Prithviraj AmmanabroluEMNLP 2022 · 被引用 1 次
- Fusing Pre-Trained Language Models with Multimodal Prompts through Reinforcement LearningYoungjae Yu, Jiwan Chung, Heeseung Yun, Jack Hessel 等CVPR 2023
它引用的顶会 Paper13
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Leveraging Procedural Generation to Benchmark Reinforcement LearningKarl Cobbe, Christopher Hesse, Jacob Hilton, John SchulmanICML 2020 · 被引用 685 次
- Emergent Complexity and Zero-shot Transfer via Unsupervised Environment DesignMichael Dennis, Natasha Jaques, Eugene Vinitsky, Alexandre M. Bayen 等NeurIPS 2020 · 被引用 362 次
- Poly-encoders: Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence ScoringSamuel Humeau, Kurt Shuster, Marie-Anne Lachaux, Jason WestonICLR 2020 · 被引用 316 次
- Interactive Fiction Games: A Colossal AdventureMatthew J. Hausknecht, Prithviraj Ammanabrolu, Marc-Alexandre Côté, Xingdi YuanAAAI 2020 · 被引用 242 次
相关 Paper
- Generating Interactive Worlds with TextAngela Fan, Jack Urbanek, Pratik Ringshia, Emily Dinan 等AAAI 2020 · 被引用 30 次
- Queens are Powerful too: Mitigating Gender Bias in Dialogue GenerationEmily Dinan, Angela Fan, Adina Williams, Jack Urbanek 等EMNLP 2020 · 被引用 14 次
- Learning Knowledge Graph-based World Models of Textual EnvironmentsPrithviraj Ammanabrolu, Mark O. RiedlNeurIPS 2021 · 被引用 43 次
- Automated curriculum generation through setter-solver interactionsSébastien Racanière, Andrew K. Lampinen, Adam Santoro, David P. Reichert 等ICLR 2020 · 被引用 41 次
- Personalized Quest and Dialogue Generation in Role-Playing Games: A Knowledge Graph- and Language Model-based ApproachTrevor Ashby, Braden K. Webb, Gregory Knapp, Jackson Searle 等CHI 2023 · 被引用 64 次
