Learning Object-Oriented Dynamics for Planning from Text
Guiliang Liu, Ashutosh Adhikari, Amir-massoud Farahmand, Pascal Poupart
摘要
The advancement of dynamics models enables model-based planning in complex environments. Dynamics models mostly study image-based games with fully observable states. Generalizing these models to Text-Based Games (TBGs), which often include partially observable states with noisy text observations, is challenging. In this work, we propose an Object-Oriented Text Dynamics (OOTD) model that enables planning algorithms to solve decision-making problems in text domains. OOTD predicts a memory graph that dynamically remembers the history of object observations and filters object-irrelevant information. To improve the robustness of dynamics, our OOTD model identifies the objects influenced by input actions and predicts beliefs of object states with independently parameterized transition layers. We develop variational objectives under the object-supervised and self-supervised settings to model the stochasticity of predicted dynamics. Empirical results show that our OOTD-based planner significantly outperforms model-free baselines in terms of sample efficiency and running scores.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Large Language Models Are Neurosymbolic ReasonersMeng Fang, Shilong Deng, Yudi Zhang, Zijing Shi 等AAAI 2024 · 被引用 53 次
- PDSketch: Integrated Domain Programming, Learning, and PlanningJiayuan Mao, Tomás Lozano-Pérez, Josh Tenenbaum, Leslie Pack KaelblingNeurIPS 2022 · 被引用 40 次
- Learning to Follow Instructions in Text-Based GamesMathieu Tuli, Andrew C. Li, Pashootan Vaezipoor, Toryn Q. Klassen 等NeurIPS 2022 · 被引用 21 次
它引用的顶会 Paper8
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski 等ICLR 2020 · 被引用 969 次
- Contrastive Learning of Structured World ModelsThomas N. Kipf, Elise van der Pol, Max WellingICLR 2020 · 被引用 322 次
- Interactive Fiction Games: A Colossal AdventureMatthew J. Hausknecht, Prithviraj Ammanabrolu, Marc-Alexandre Côté, Xingdi YuanAAAI 2020 · 被引用 242 次
- Graph Constrained Reinforcement Learning for Natural Language Action SpacesPrithviraj Ammanabrolu, Matthew J. HausknechtICLR 2020 · 被引用 138 次
- Learning Dynamic Belief Graphs to Generalize on Text-Based GamesAshutosh Adhikari, Xingdi Yuan, Marc-Alexandre Côté, Mikulas Zelinka 等NeurIPS 2020 · 被引用 91 次
相关 Paper
- Dyn-O: Building Structured World Models with Object-Centric RepresentationsZizhao Wang, Kaixin Wang, Li Zhao, Peter Stone 等NeurIPS 2025 · 被引用 15 次
- Object-Oriented Dynamics Learning through Multi-Level AbstractionGuangxiang Zhu, Jianhao Wang, Zhizhou Ren, Zichuan Lin 等AAAI 2020 · 被引用 6 次
- Eye of the Beholder: Improved Relation Generalization for Text-Based Reinforcement Learning AgentsKeerthiram Murugesan, Subhajit Chaudhury, Kartik TalamadupulaAAAI 2022 · 被引用 5 次
- Learning Knowledge Graph-based World Models of Textual EnvironmentsPrithviraj Ammanabrolu, Mark O. RiedlNeurIPS 2021 · 被引用 43 次
- Working Memory GraphsRicky Loynd, Roland Fernandez, Asli Celikyilmaz, Adith Swaminathan 等ICML 2020 · 被引用 41 次
