People Do Not Just Plan, They Plan to Plan
Mark K. Ho, David Abel, Jonathan D. Cohen, Michael L. Littman, Thomas L. Griffiths
Abstract
Planning is useful. It lets people take actions that have desirable long-term consequences. But, planning is hard. It requires thinking about consequences, which consumes limited computational and cognitive resources. Thus, people should plan their actions, but they should also be smart about how they deploy resources used for planning their actions. Put another way, people should also “plan their plans”. Here, we formulate this aspect of planning as a meta-reasoning problem and formalize it in terms of a recursive Bellman objective that incorporates both task rewards and information-theoretic planning costs. Our account makes quantitative predictions about how people should plan and meta-plan as a function of the overall structure of a task, which we test in two experiments with human participants. We find that people's reaction times reflect a planned use of information processing, consistent with our account. This formulation of planning to plan provides new insight into the function of hierarchical planning, state abstraction, and cognitive control in both humans and machines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 06318c85-b0ce-489f-99fe-165552b5d7ddRelated papers
- Stop! Planner Time: Metareasoning for Probabilistic Planning Using Learned Performance ProfilesMatthew Budd, Bruno Lacerda, Nick HawesAAAI 2024 · 2 citations
- Modeling Boundedly Rational Agents with Latent Inference BudgetsAthul Paul Jacob, Abhishek Gupta, Jacob AndreasICLR 2024 · 4 citations
- Forethought and Hindsight in Credit AssignmentVeronica Chelu, Doina Precup, Hado van HasseltNeurIPS 2020 · 29 citations
- Dynamic allocation of limited memory resources in reinforcement learningNisheet Patel, Luigi Acerbi, Alexandre PougetNeurIPS 2020 · 6 citations
- Inferring Rewards from Language in ContextJessy Lin, Daniel Fried, Dan Klein, Anca D. DraganACL 2022 · 71 citations
