VECA: A New Benchmark and Toolkit for General Cognitive Development
Kwanyoung Park, Hyunseok Oh, Youngki Lee
摘要
The developmental approach, simulating a cognitive development of a human, arises as a way to nurture a human-level commonsense and overcome the limitations of data-driven approaches. However, neither a virtual environment nor an evaluation platform exists for the overall development of core cognitive skills. We present the VECA(Virtual Environment for Cognitive Assessment), which consists of two main components: (i) a first benchmark to assess the overall cognitive development of an AI agent, and (ii) a novel toolkit to generate diverse and distinct cognitive tasks. VECA benchmark virtually implements the cognitive scale of Bayley Scales of Infant and Toddler Development-IV(Bayley-4), the gold-standard developmental assessment for human infants and toddlers. Our VECA toolkit provides a human toddler-like embodied agent with various human-like perceptual features crucial to human cognitive development, e.g., binocular vision, 3D-spatial audio, and tactile receptors. We compare several modern RL algorithms on our VECA benchmark and seek their limitations in modeling human-like cognitive development. We further analyze the validity of the VECA benchmark, as well as the effect of human-like sensory characteristics on cognitive skills.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- AGENT: A Benchmark for Core Psychological ReasoningTianmin Shu, Abhishek Bhandwaldar, Chuang Gan, Kevin A. Smith 等ICML 2021 · 被引用 79 次
- CURI: A Benchmark for Productive Concept Learning Under UncertaintyRamakrishna Vedantam, Arthur Szlam, Maximilian Nickel, Ari Morcos 等ICML 2021 · 被引用 32 次
- Elastic Tactile Simulation Towards Tactile-Visual PerceptionYikai Wang, Wenbing Huang, Bin Fang, Fuchun Sun 等ACM MM 2021 · 被引用 20 次
相关 Paper
- BabyVLM-V2: Toward Developmentally Grounded Pretraining and Benchmarking of Vision Foundation ModelsShengao Wang, Wenqi Wang, Zecheng Wang, Max Whitton 等CVPR 2026 · 被引用 4 次
- Easy for Children, Hard for AI: The Limits of Multimodal LLMs in Early Childhood LearningJingping Liu, Xueyan Wu, Hanxuan Chen, Ziyan Liu 等AAAI 2026
- ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition BenchmarkRonghao Dang, Yuqian Yuan, Wenqi Zhang, Yifei Xin 等CVPR 2025
- Lookee: Gaze Tracking-based Infant Vocabulary Comprehension Assessment and AnalysisMinji Kim, Minkyu Shim, Jun Ho Chai, Eon-Suk Ko 等CHI 2025 · 被引用 2 次
- VirtualEnv: A Platform for Embodied AI ResearchKabir Swain, Sijie Han, Ayush Raina, Jin Zhang 等AAAI 2026
