MLE-STAR: Machine Learning Engineering Agent via Search and Targeted Refinement
Jaehyun Nam, Jinsung Yoon, Jiefeng Chen, Jinwoo Shin, Sercan Ö. Arik, Tomas Pfister
摘要
Agents based on large language models (LLMs) for machine learning engineering (MLE) can automatically implement ML models via code generation. However, existing approaches to build such agents often rely heavily on inherent LLM knowledge and employ coarse exploration strategies that modify the entire code structure at once. This limits their ability to select effective task-specific models and perform deep exploration within specific components, such as experimenting extensively with feature engineering options. To overcome these, we propose MLE-STAR, a novel approach to build MLE agents. MLE-STAR first leverages external knowledge by using a search engine to retrieve effective models from the web, forming an initial solution, then iteratively refines it by exploring various strategies targeting specific ML components. This exploration is guided by ablation studies analyzing the impact of individual code blocks. Furthermore, we introduce a novel ensembling method using an effective strategy suggested by MLE-STAR. Our experimental results show that MLE-STAR achieves medals in 64% of the Kaggle competitions on the MLE-bench, significantly outperforming the best alternative. 1 * This work was done while Jaehyun was a student researcher at Google Cloud. 1 We release open-source codebase of MLE-STAR at https://github.com/google/adk-samples.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- MARS: Modular Agent with Reflective Search for Automated AI ResearchJiefeng Chen, Bhavana Dalvi Mishra, Jaehyun Nam, Rui Meng 等ICML 2026 · 被引用 14 次
- Can We Predict Before Executing Machine Learning Agents?Jingsheng Zheng, Jintian Zhang, Yujie Luo, Yuren Mao 等ACL 2026 · 被引用 6 次
- TusoAI: Agentic Optimization for Scientific MethodsAlistair Turcan, Kexin Huang, Lei Li, Martin J. ZhangICLR 2026 · 被引用 3 次
- EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context ManagementZherui Yang, Fan Liu, Yansong Ning, Hao LiuKDD 2026 · 被引用 3 次
- stratum: A System Infrastructure for Massive Agent-Centric ML WorkloadsArnab Phani, Elias Strauss, Sebastian SchelterVLDB 2026
它引用的顶会 Paper14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging FaceYongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li 等NeurIPS 2023 · 被引用 1,778 次
- Large Language Models for Automated Data Science: Introducing CAAFE for Context-Aware Automated Feature EngineeringNoah Hollmann, Samuel Müller, Frank HutterNeurIPS 2023 · 被引用 210 次
- MLAgentBench: Evaluating Language Agents on Machine Learning ExperimentationQian Huang, Jian Vora, Percy Liang, Jure LeskovecICML 2024 · 被引用 209 次
- Better by default: Strong pre-tuned MLPs and boosted trees on tabular dataDavid Holzmüller, Léo Grinsztajn, Ingo SteinwartNeurIPS 2024 · 被引用 141 次
相关 Paper
- CoMind: Towards Community-Driven Agents for Machine Learning EngineeringSijie Li, Weiwei Sun, Shanda Li, Ameet Talwalkar 等ICLR 2026 · 被引用 3 次
- Investigating Component Contributions in Multi-Agent ML SystemsJunsung Kim, Ilia Mireskandari, Seungwan Son, Yifan Zhou 等ICML 2026
- DS-Agent: Automated Data Science by Empowering Large Language Models with Case-Based ReasoningSiyuan Guo, Cheng Deng, Ying Wen, Hechang Chen 等ICML 2024 · 被引用 107 次
- Adaptive Self-improvement LLM Agentic System for ML Library DevelopmentGenghan Zhang, Weixin Liang, Olivia Hsu, Kunle OlukotunICML 2025
- Improving Knowledge Extraction from LLMs for Task Learning through Agent AnalysisJames R. Kirk, Robert E. Wray, Peter Lindes, John E. LairdAAAI 2024 · 被引用 8 次
