AutoToM: Scaling Model-based Mental Inference via Automated Agent Modeling
Zhining Zhang, Chuanyang Jin, Mung Yao Jia, Shunchi Zhang, Tianmin Shu
Abstract
Theory of Mind (ToM), the ability to understand people's minds based on their behavior, is key to developing socially intelligent agents. Current approaches to ToM reasoning either rely on prompting Large Language Models (LLMs), which are prone to systematic errors, or use handcrafted, rigid agent models for model-based inference, which are more robust but fail to generalize across domains. In this work, we introduce AutoToM , an automated agent modeling method for scalable, robust, and interpretable mental inference. Given a ToM problem, AutoToM first proposes an initial agent model and then performs automated Bayesian inverse planning based on this model, leveraging an LLM backend. Guided by inference uncertainty, it iteratively refines the model by introducing additional mental variables and/or incorporating more timesteps in the context. Across five diverse benchmarks, AutoToM outperforms existing ToM methods and even large reasoning models. Additionally, we show that AutoToM can produce human-like confidence estimates and enable online mental inference for embodied decision-making.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 56789746-db3f-43fb-9b41-982fee046076Cited by top-tier papers6
- ToMAP: Training Opponent-Aware LLM Persuaders with Theory of MindPeixuan Han, Zijia Liu, Jiaxuan YouICML 2026 · 9 citations
- MindZero: Learning Online Mental Reasoning With Zero AnnotationsShunchi Zhang, Jin Lu, Chuanyang Jin, Yichao Zhou et al.ICML 2026 · 1 citation
- TactfulToM: Do LLMs have the Theory of Mind ability to understand White Lies?Yiwei Liu, Emma Jane Pretty, Jiahao Huang, Saku SugawaraEMNLP 2025 · 1 citation
- Reality vs Counterfactual: Multi-World Contrastive Reinforcement Learning for Enhancing MLLM's Theory of Mind in Egocentric VideosGuiyang Hou, Yihui Fu, Chen Wu, Xiang Huang et al.AAAI 2026
- From Shortcuts to Reasoning: Robust Post-Training of Theory of Mind with Reinforcement LearningJike Zhong, Yuxiang Lai, Ming Li, Yuheng Li et al.ICML 2026
Builds on18
- Hypothesis Search: Inductive Reasoning with Language ModelsRuocheng Wang, Eric Zelikman, Gabriel Poesia, Yewen Pu et al.ICLR 2024 · 156 citations
- Online Bayesian Goal Inference for Boundedly Rational Planning AgentsTan Zhi-Xuan, Jordyn L. Mann, Tom Silver, Josh Tenenbaum et al.NeurIPS 2020 · 122 citations
- Towards Mutual Theory of Mind in Human-AI Interaction: How Language Reflects What Students Perceive About a Virtual Teaching AssistantQiaosi Wang, Koustuv Saha, Eric Gregori, David A. Joyner et al.CHI 2021 · 115 citations
- Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis RefinementLinlu Qiu, Liwei Jiang, Ximing Lu, Melanie Sclar et al.ICLR 2024 · 114 citations
- Neural Theory-of-Mind? On the Limits of Social Intelligence in Large LMsMaarten Sap, Ronan Le Bras, Daniel Fried, Yejin ChoiEMNLP 2022 · 92 citations
Related papers
- Overcoming Multi-step Complexity in Multimodal Theory-of-Mind Reasoning: A Scalable Bayesian PlannerChunhui Zhang, Zhongyu Ouyang, Kwonjoon Lee, Nakul Agarwal et al.ICML 2025
- Tracing Belief-Driven Thoughts with Theory-of-Mind Agents: An Opinion Analysis FrameworkJintao Wen, Yunfeng Ning, Hankun Kang, Xin Miao et al.WWW 2026
- MuMA-ToM: Multi-modal Multi-Agent Theory of MindHaojun Shi, Suyu Ye, Xinyu Fang, Chuanyang Jin et al.AAAI 2025 · 48 citations
- MetaMind: Modeling Human Social Thoughts with Metacognitive Multi-Agent SystemsXuanming Zhang, Yuxuan Chen, Samuel (Min-Hsuan) Yeh, Sharon LiNeurIPS 2025 · 14 citations
- MMToM-QA: Multimodal Theory of Mind Question AnsweringChuanyang Jin, Yutong Wu, Jing Cao, Jiannan Xiang et al.ACL 2024 · 8 citations
