Learning to execute instructions in a Minecraft dialogue
Prashant Jayannavar, Anjali Narayan-Chen, Julia Hockenmaier
2020Year
22Citations
6Top-tier citations
Abstract
The Minecraft Collaborative Building Task is a two-player game in which an Architect A instructs a Builder B to construct a target structure out of 3D blocks. We consider the task of predicting B's action sequences (block placements and removals) in a given game context, and show that capturing B's past actions as well as B's perspective leads to a significant improvement in performance on this challenging language understanding problem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext af514676-7d38-45f4-a711-26ead816f69aCited by top-tier papers6
- MindCraft: Theory of Mind Modeling for Situated Dialogue in Collaborative TasksCristian-Paul Bara, Sky CH-Wang, Joyce ChaiEMNLP 2021 · 28 citations
- Chain of Execution Supervision Promotes General Reasoning in Large Language ModelsNuo Chen, Zehua Li, Keqin Bao, Junyang Lin et al.NeurIPS 2025 · 6 citations
- Gesturing Toward Abstraction: Multimodal Convention Formation in Collaborative Physical TasksKiyosu Maeda, William P. McCarthy, Ching-Yi Tsai, Jeffrey Mu et al.CHI 2026 · 1 citation
- Glider: Global and Local Instruction-Driven Expert RouterPingzhi Li, Prateek Yadav, Jaehong Yoon, Jie Peng et al.EMNLP 2025 · 1 citation
- MindZero: Learning Online Mental Reasoning With Zero AnnotationsShunchi Zhang, Jin Lu, Chuanyang Jin, Yichao Zhou et al.ICML 2026 · 1 citation
Related papers
- CraftAssist Instruction Parsing: Semantic Parsing for a Voxel-World AssistantKavya Srinet, Yacine Jernite, Jonathan Gray, Arthur SzlamACL 2020 · 7 citations
- Optimus-2: Multimodal Minecraft Agent with Goal-Observation-Action Conditioned PolicyZaijing Li, Yuquan Xie, Rui Shao, Gongwei Chen et al.CVPR 2025
- OmniJARVIS: Unified Vision-Language-Action Tokenization Enables Open-World Instruction Following AgentsZihao Wang, Shaofei Cai, Zhancun Mu, Haowei Lin et al.NeurIPS 2024 · 37 citations
- Order-Aware Generative Modeling Using the 3D-Craft DatasetZhuoyuan Chen, Kavya Srinet, Charles R. Qi, Haoqi Fan et al.ICCV 2019 · 8 citations
- Program Guided AgentShao-Hua Sun, Te-Lin Wu, Joseph J. LimICLR 2020 · 63 citations
