Perceiving the World: Question-guided Reinforcement Learning for Text-based Games
Yunqiu Xu, Meng Fang, Ling Chen, Yali Du, Joey Tianyi Zhou, Chengqi Zhang
Abstract
Text-based games provide an interactive way to study natural language processing. While deep reinforcement learning has shown effectiveness in developing the game playing agent, the low sample efficiency and the large action space remain to be the two major challenges that hinder the DRL from being applied in the real world. In this paper, we address the challenges by introducing world-perceiving modules, which automatically decompose tasks and prune actions by answering questions about the environment. We then propose a two-phase training framework to decouple language learning from reinforcement learning, which further improves the sample efficiency. The experimental results show that the proposed method significantly improves the performance and sample efficiency. Besides, it shows robustness against compound error and limited pre-training data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 07e95986-c159-41f2-aa61-3740b840d699Cited by top-tier papers5
- Large Language Models Are Neurosymbolic ReasonersMeng Fang, Shilong Deng, Yudi Zhang, Zijing Shi et al.AAAI 2024 · 53 citations
- EAGER: Asking and Answering Questions for Automatic Reward Shaping in Language-guided RLThomas Carta, Pierre-Yves Oudeyer, Olivier Sigaud, Sylvain LamprierNeurIPS 2022 · 35 citations
- Persona Dynamics: Unveiling the Impact of Persona Traits on Agents in Text-Based GamesSeungwon Lim, Seungbeen Lee, Dongjun Min, Youngjae YuACL 2025 · 1 citation
- Monte Carlo Planning with Large Language Model for Text-Based Game AgentsZijing Shi, Meng Fang, Ling ChenICLR 2025
- Language Model Adaption for Reinforcement Learning with Natural Language Action SpaceJiangxing Wang, Jiachen Li, Xiao Han, Deheng Ye et al.ACL 2024
Builds on24
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
- Data Augmentation for Graph Neural NetworksTong Zhao, Yozen Liu, Leonardo Neves, Oliver J. Woodford et al.AAAI 2021 · 487 citations
- Dynamics-Aware Unsupervised Discovery of SkillsArchit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar et al.ICLR 2020 · 475 citations
- Skew-Fit: State-Covering Self-Supervised Reinforcement LearningVitchyr Pong, Murtaza Dalal, Steven Lin, Ashvin Nair et al.ICML 2020 · 303 citations
- SQIL: Imitation Learning via Reinforcement Learning with Sparse RewardsSiddharth Reddy, Anca D. Dragan, Sergey LevineICLR 2020 · 299 citations
Related papers
- PLM-based World Models for Text-based GamesMinsoo Kim, YeonJoon Jung, Dohyeon Lee, Seung-won HwangEMNLP 2022 · 4 citations
- Learning Knowledge Graph-based World Models of Textual EnvironmentsPrithviraj Ammanabrolu, Mark O. RiedlNeurIPS 2021 · 43 citations
- LeDeepChef Deep Reinforcement Learning Agent for Families of Text-Based GamesLeonard Adolphs, Thomas HofmannAAAI 2020 · 48 citations
- Text-based RL Agents with Commonsense Knowledge: New Challenges, Environments and BaselinesKeerthiram Murugesan, Mattia Atzeni, Pavan Kapanipathi, Pushkar Shukla et al.AAAI 2021 · 60 citations
- Learning to Follow Instructions in Text-Based GamesMathieu Tuli, Andrew C. Li, Pashootan Vaezipoor, Toryn Q. Klassen et al.NeurIPS 2022 · 21 citations
