Baby Intuitions Benchmark (BIB): Discerning the goals, preferences, and actions of others
Kanishk Gandhi, Gala Stojnic, Brenden M. Lake, Moira R. Dillon
摘要
To achieve human-like common sense about everyday life, machine learning systems must understand and reason about the goals, preferences, and actions of other agents in the environment. By the end of their first year of life, human infants intuitively achieve such common sense, and these cognitive achievements lay the foundation for humans' rich and complex understanding of the mental states of others. Can machines achieve generalizable, commonsense reasoning about other agents like human infants? The Baby Intuitions Benchmark (BIB) 1 challenges machines to predict the plausibility of an agent's behavior based on the underlying causes of its actions. Because BIB's content and paradigm are adopted from developmental cognitive science, BIB allows for direct comparison between human and machine performance. Nevertheless, recently proposed, deep-learning-based agency reasoning models fail to show infant-like reasoning, leaving BIB an open challenge.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- ToM2C: Target-oriented Multi-agent Communication and Cooperation with Theory of MindYuanfei Wang, Fangwei Zhong, Jing Xu, Yizhou WangICLR 2022 · 被引用 103 次
- AGENT: A Benchmark for Core Psychological ReasoningTianmin Shu, Abhishek Bhandwaldar, Chuang Gan, Kevin A. Smith 等ICML 2021 · 被引用 79 次
- MuMA-ToM: Multi-modal Multi-Agent Theory of MindHaojun Shi, Suyu Ye, Xinyu Fang, Chuanyang Jin 等AAAI 2025 · 被引用 48 次
- Language Models Represent Beliefs of Self and OthersWentao Zhu, Zhining Zhang, Yizhou WangICML 2024 · 被引用 24 次
- Symmetric Machine Theory of MindMelanie Sclar, Graham Neubig, Yonatan BiskICML 2022 · 被引用 22 次
它引用的顶会 Paper3
- Decoupling Representation Learning from Reinforcement LearningAdam Stooke, Kimin Lee, Pieter Abbeel, Michael LaskinICML 2021 · 被引用 389 次
- Keep Doing What Worked: Behavior Modelling Priors for Offline Reinforcement LearningNoah Y. Siegel, Jost Tobias Springenberg, Felix Berkenkamp, Abbas Abdolmaleki 等ICLR 2020 · 被引用 299 次
- AGENT: A Benchmark for Core Psychological ReasoningTianmin Shu, Abhishek Bhandwaldar, Chuang Gan, Kevin A. Smith 等ICML 2021 · 被引用 79 次
相关 Paper
- Neural Reasoning about Agents' Goals, Preferences, and ActionsMatteo Bortoletto, Lei Shi, Andreas BullingAAAI 2024 · 被引用 8 次
- BDIQA: A New Dataset for Video Question Answering to Explore Cognitive Reasoning through Theory of MindYuanyuan Mao, Xin Lin, Qin Ni, Liang HeAAAI 2024 · 被引用 6 次
- X-VoE: Measuring eXplanatory Violation of Expectation in Physical EventsBo Dai, Linge Wang, Baoxiong Jia, Zeyu Zhang 等ICCV 2023 · 被引用 4 次
- VECA: A New Benchmark and Toolkit for General Cognitive DevelopmentKwanyoung Park, Hyunseok Oh, Youngki LeeAAAI 2022
- A Newborn Embodied Turing Test for Comparing Object Segmentation Across Animals and MachinesManju Garimella, Denizhan Pak, Justin N. Wood, Samantha Marie Waters WoodICLR 2024 · 被引用 1 次
