Probing Emergent Semantics in Predictive Agents via Question Answering
Abhishek Das, Federico Carnevale, Hamza Merzic, Laura Rimell, Rosalia Schneider, Josh Abramson, Alden Hung, Arun Ahuja, Stephen Clark, Greg Wayne, Felix Hill
摘要
Recent work has shown how predictive modeling can endow agents with rich knowledge of their surroundings, improving their ability to act in complex environments. We propose question-answering as a general paradigm to decode and understand the representations that such agents develop, applying our method to two recent approaches to predictive modeling -action-conditional CPC (Guo et al., 2018) and SimCore (Gregor et al., 2019). After training agents with these predictive objectives in a visually-rich, 3D environment with an assortment of objects, colors, shapes, and spatial configurations, we probe their internal state representations with synthetic (English) questions, without backpropagating gradients from the question-answering decoder into the agent. The performance of different agents when probed this way reveals that they learn to encode factual, and seemingly compositional, information about objects, properties and spatial relations from their physical environment. Our approach is intuitive, i.e. humans can easily interpret responses of the model as opposed to inspecting continuous vectors, and model-agnostic, i.e. applicable to any modeling approach. By revealing the implicit knowledge of objects, quantities, properties and relations acquired by agents as they learn, question-conditional agent probing can stimulate the design and development of stronger predictive learning objectives.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Auxiliary Tasks and Exploration Enable ObjectGoal NavigationJoel Ye, Dhruv Batra, Abhishek Das, Erik WijmansICCV 2021 · 被引用 137 次
- Habitat-Web: Learning Embodied Object-Search Strategies from Human Demonstrations at ScaleRam Ramrakhya, Eric Undersander, Dhruv Batra, Abhishek DasCVPR 2022 · 被引用 73 次
- VLM Agents Generate Their Own Memories: Distilling Experience into Embodied Programs of ThoughtGabriel Sarch, Lawrence Jang, Michael J. Tarr, William W. Cohen 等NeurIPS 2024 · 被引用 64 次
- Perceiving the World: Question-guided Reinforcement Learning for Text-based GamesYunqiu Xu, Meng Fang, Ling Chen, Yali Du 等ACL 2022 · 被引用 22 次
- EXCALIBUR: Encouraging and Evaluating Embodied ExplorationHao Zhu, Raghav Kapoor, So Yeon Min, Winson Han 等CVPR 2023
相关 Paper
- Environment Predictive Coding for Visual NavigationSanthosh Kumar Ramakrishnan, Tushar Nagarajan, Ziad Al-Halah, Kristen GraumanICLR 2022 · 被引用 10 次
- Capturing Visual Environment Structure Correlates with Control PerformanceJiahua Dong, Yunze Man, Pavel Tokmakov, Yu-Xiong WangICLR 2026 · 被引用 2 次
- Causal-JEPA: Learning World Models through Object-Level Latent MaskingHeejeong Nam, Quentin Le Lidec, Lucas Maes, Yann LeCun 等ICML 2026 · 被引用 7 次
- Predict Before You Explore: Predictive Planning with Specialized Memory for Embodied Question AnsweringBowen Yuan, Sisi You, Bing-Kun BaoCVPR 2026
- Agent Modelling under Partial Observability for Deep Reinforcement LearningGeorgios Papoudakis, Filippos Christianos, Stefano V. AlbrechtNeurIPS 2021 · 被引用 110 次
