Cradle: Empowering Foundation Agents towards General Computer Control
Weihao Tan, Wentao Zhang, Xinrun Xu, Haochong Xia, Ziluo Ding, Boyu Li, Bohan Zhou, Junpeng Yue, Jiechuan Jiang, Yewen Li, Ruyi An, Molei Qin
Abstract
agents.github.io/Cradle Everyday I n f o r m a t i o n G a t h e r i n g A c t i o n P l a n n i n g T a s k I n f e r e n c e S k i l l C u r a t i o n S e l f -R e f l e c t i o n R e t r i e v e d S k i l l s : [' a i m ' ,' f o l l o w ' ,' t u r n ' ,' m o v e _ f o r w a r d ' ,'s h o w _ w e a p o n _ w h e e l ' . . . ]L a s t A c t i o n : t a k e _ c o v e r ( ) P r e v i o u s T a s k : P r e s s [ Q ] t o t a k e c o v e r . T h e l a s t a c t i o n w a s e x e c u t e d s u c c e s s f u l l y , a n d t h e t a s k i s c o m p l e t e d s i n c e n e w g u i d a n c e a p p e a r s . H o l d [ T A B ] t o s h o w t h e W e a p o n W h e e l . s h o w _ w e a p o n _ w h e e l ( ) H o l d [ T A B ] t o s h o w t h e W e a p o n W h e e l . G e n e r a t e d S k i l l : d e f s h o w _ w e a p o n _ w h e e l ( ) : i o _ e n v . k e y _ h o l d (' t a b ' )G u i d a n c e : # S e l e c t s p a r s n i p s e e d s f r o m t h e t o o l b a r . s e l e c t _ t o o l (k e y =' 6 ' ) # P l a n t s p a r s n i p s e e d s o n t h e t i l l e d s o i l . d o _ a c t i o n ( ) R e t r i e v e d S k i l l s : [' m o v e _ u p ' ,' m o v e _ d o w n ' ,' m o v e _ l e f t ' ,' m o v e _ r i g h t ' ,' s e l e c t _ t o o l ' ,' d o _ a c t i o n ' . .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 731c7327-cc5f-4449-8b53-a58f6d48765aCited by top-tier papers30
- Group-in-Group Policy Optimization for LLM Agent TrainingLang Feng, Zhenghai Xue, Tingcong Liu, Bo AnNeurIPS 2025 · 484 citations
- lmgame-Bench: How Good are LLMs at Playing Games?Lanxiang Hu, Mingjia Huo, Yuxuan Zhang, Haoyang Yu et al.ICLR 2026 · 47 citations
- Hierarchy-of-Groups Policy Optimization for Long-Horizon Agentic TasksShuo He, Lang Feng, Qi Wei, Xin Cheng et al.ICLR 2026 · 36 citations
- Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video GamesDongmin Park, Minkyu Kim, Beongjun Choi, Junhyuck Kim et al.ICLR 2026 · 30 citations
- NitroGen: An Open Foundation Model for Generalist Gaming AgentsLoïc Magne, Anas Awadalla, Guanzhi Wang, Yinzhen Xu et al.CVPR 2026 · 22 citations
Builds on23
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan et al.NeurIPS 2023 · 5,828 citations
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran et al.NeurIPS 2023 · 5,068 citations
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch et al.ICML 2023 · 2,601 citations
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris et al.UIST 2023 · 1,882 citations
Related papers
- Towards Unobtrusive Physical AI: Augmenting Everyday Objects with Intelligence and Robotic Movement for Proactive AssistanceViolet Yinuo Han, Jesse T. Gonzalez, Christina Yang, Zhiruo Wang et al.UIST 2025 · 3 citations
- MobileIPL: Enhancing Mobile Agents Thinking Process via Iterative Preference LearningHuang Kun, Weikai Xu, Yuxuan Liu, Quandong Wang et al.ICLR 2026 · 9 citations
- CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation ModelZhuoyuan Yu, Yuxing Long, Zihan Yang, Chengyan Zeng et al.AAAI 2026 · 14 citations
- Action-Sketcher: From Reasoning to Action via Visual Sketches for Robotic ManipulationHuajie Tan, Peterson Co, Yijie Xu, Shanyu Rong et al.CVPR 2026
- Learn the Ropes, Then Trust the Wins: Self-imitation with Progressive Exploration for Agentic Reinforcement LearningYulei Qin, Xiaoyu Tan, Zhengbao He, Gang Li et al.ICLR 2026 · 9 citations
