Monte-Carlo Planning and Learning with Language Action Value Estimates
Youngsoo Jang, Seokin Seo, Jongmin Lee, Kee-Eung Kim
摘要
Interactive Fiction (IF) games provide a useful testbed for language-based reinforcement learning agents, posing significant challenges of natural language understanding, commonsense reasoning, and non-myopic planning in the combinatorial search space. Agents based on standard planning algorithms struggle to play IF games due to the massive search space of language actions. Thus, language-grounded planning is a key ability of such agents, since inferring the consequence of language action based on semantic understanding can drastically improve search. In this paper, we introduce Monte-Carlo planning with Language Action Value Estimates (MC-LAVE) that combines a Monte-Carlo tree search with language-driven exploration. MC-LAVE invests more search effort into semantically promising language actions using locally optimistic language value estimates, yielding a significant reduction in the effective search space of language actions. We then present a reinforcement learning approach via MC-LAVE, which alternates between MC-LAVE planning and supervised learning of the self-generated language actions. In the experiments, we demonstrate that our method achieves new high scores in various IF games.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper10
- WebPilot: A Versatile and Autonomous Multi-Agent System for Web Task Execution with Strategic ExplorationYao Zhang, Zijian Ma, Yunpu Ma, Zhen Han 等AAAI 2025 · 被引用 101 次
- Multi-Stage Episodic Control for Strategic Exploration in Text GamesJens Tuyls, Shunyu Yao, Sham M. Kakade, Karthik NarasimhanICLR 2022 · 被引用 30 次
- Model-Based Simulation for Optimising Smart ReplyBenjamin Towle, Ke ZhouACL 2023 · 被引用 1 次
- Broaden your SCOPE! Efficient Multi-turn Conversation Planning for LLMs with Semantic SpaceZhiliang Chen, Xinyuan Niu, Chuan-Sheng Foo, Bryan Kian Hsiang LowICLR 2025
- Monte Carlo Planning with Large Language Model for Text-Based Game AgentsZijing Shi, Meng Fang, Ling ChenICLR 2025
相关 Paper
- Graph Constrained Reinforcement Learning for Natural Language Action SpacesPrithviraj Ammanabrolu, Matthew J. HausknechtICLR 2020 · 被引用 138 次
- Interactive Fiction Games: A Colossal AdventureMatthew J. Hausknecht, Prithviraj Ammanabrolu, Marc-Alexandre Côté, Xingdi YuanAAAI 2020 · 被引用 242 次
- Interactive Fiction Game Playing as Multi-Paragraph Reading Comprehension with Reinforcement LearningXiaoxiao Guo, Mo Yu, Yupeng Gao, Chuang Gan 等EMNLP 2020 · 被引用 21 次
- Text-based RL Agents with Commonsense Knowledge: New Challenges, Environments and BaselinesKeerthiram Murugesan, Mattia Atzeni, Pavan Kapanipathi, Pushkar Shukla 等AAAI 2021 · 被引用 60 次
- Deep Reinforcement Learning with Stacked Hierarchical Attention for Text-based GamesYunqiu Xu, Meng Fang, Ling Chen, Yali Du 等NeurIPS 2020 · 被引用 48 次
