Navigates Like Me: Understanding How People Evaluate Human-Like AI in Video Games
Stephanie Milani, Arthur Juliani, Ida Momennejad, Raluca Georgescu, Jaroslaw Rzepecki, Alison Shaw, Gavin Costello, Fei Fang, Sam Devlin, Katja Hofmann
Abstract
We aim to understand how people assess human likeness in navigation produced by people and artificially intelligent (AI) agents in a video game. To this end, we propose a novel AI agent with the goal of generating more human-like behavior. We collect hundreds of crowd-sourced assessments comparing the human-likeness of navigation behavior generated by our agent and baseline AI agents with human-generated behavior. Our proposed agent passes a Turing Test, while the baseline agents do not. By passing a Turing Test, we mean that human judges could not quantitatively distinguish between videos of a person and an AI agent navigating. To understand what people believe constitutes human-like navigation, we extensively analyze the justifications of these assessments. This work provides insights into the characteristics that people consider human-like in the context of goal-directed video game navigation, which is a key step for further improving human interactions with AI agents.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Learning Human-Like RL Agents Through Trajectory Optimization With Action QuantizationJian-Ting Guo, Yu-Cheng Chen, Ping-Chun Hsieh, Kuo-Hao Ho et al.NeurIPS 2025 · 3 citations
- Playing the Imitation Game: How Perceived Generated Content Shapes Player ExperienceMahsa Bazzaz, Seth CooperCHI 2026
Builds on5
- Implementation Matters in Deep RL: A Case Study on PPO and TRPOLogan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras et al.ICLR 2020 · 305 citations
- Multi-Game Decision TransformersKuang-Huei Lee, Ofir Nachum, Mengjiao Yang, Lisa Lee et al.NeurIPS 2022 · 279 citations
- Modeling Strong and Human-Like Gameplay with KL-Regularized SearchAthul Paul Jacob, David J. Wu, Gabriele Farina, Adam Lerer et al.ICML 2022 · 69 citations
- Navigation Turing Test (NTT): Learning to Evaluate Human-Like NavigationSam Devlin, Raluca Georgescu, Ida Momennejad, Jaroslaw Rzepecki et al.ICML 2021 · 27 citations
- Emergence of Maps in the Memories of Blind Navigation AgentsErik Wijmans, Manolis Savva, Irfan Essa, Stefan Lee et al.ICLR 2023 · 17 citations
Related papers
- Human or Machine? A Preliminary Turing Test for Speech-to-Speech InteractionXiang Li, Jiabao Gao, Sipei Lin, Xuan Zhou et al.ICLR 2026 · 1 citation
- A Computational Framework for Evaluating Human-likeness in LLMs' Open-ended Human BehaviorsYuxuan Lei, Jianxun Lian, Defu Lian, Jincenzi Wu et al.ICML 2026
- Towards Motion Turing Test: Evaluating Human-Likeness in Humanoid RobotsMingzhe Li, Mengyin Liu, Zekai Wu, Xincheng Lin et al.CVPR 2026 · 4 citations
- Effects of Communication Directionality and AI Agent Differences in Human-AI InteractionZahra Ashktorab, Casey Dugan, James Johnson, Qian Pan et al.CHI 2021 · 47 citations
- Not All the Same: Understanding and Informing Similarity Estimation in Tile-Based Video GamesSebastian Berns, Vanessa Volz, Laurissa Tokarchuk, Sam Snodgrass et al.CHI 2024 · 2 citations
