AgentExpt: Automating AI Experiment Design with LLM-based Resource Retrieval Agent
Yu Li, Lehui Li, Lin Chen, Qingmin Liao, Fengli Xu, Yong Li
摘要
In modern AI research, baseline and dataset selection is a high-stakes decision in experimental design. It operationalizes a research idea into a concrete evaluation protocol and largely determines the validity and comparability of empirical conclusions. However, making appropriate choices is increasingly difficult as baselines and datasets proliferate, while suitability is inherently context-dependent and rarely captured by baseline and dataset metadata. To address these challenges, we present AgentExpt, a comprehensive framework for baseline and dataset recommendation. We first curate a large-scale, high-quality knowledge base that links 108,825 accepted papers to their used baselines and datasets. Based on this resource, we design a collective perception-enhanced retriever that represents each baseline or dataset by integrating first-person self-descriptions with third-person citation contexts, thereby effectively positioning them within the scholarly network. We further design a reasoning-augmented reranker that encodes baseline-dataset interaction chains as a reasoning prior to fine-tune an LLM, producing refined rankings with interpretable justifications. Experiments show that our framework outperforms the strongest baseline, with average gains of +5.85% in Recall@20 and +7.90% in HitRate@10, and ablation studies confirm the effectiveness of our designed components. Overall, AgentExpt advances the efficient and reliable automation of experimental design. Our code is available at https://anonymous.4open.science/r/Agentexpt-DD3E.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- AI-Researcher: Autonomous Scientific InnovationJiabin Tang, Lianghao Xia, Zhonghang Li, Chao HuangNeurIPS 2025 · 被引用 101 次
- IdeaSynth: Iterative Research Idea Development Through Evolving and Composing Idea Facets with Literature-Grounded FeedbackKevin Pu, K. J. Kevin Feng, Tovi Grossman, Tom Hope 等CHI 2025 · 被引用 15 次
- DataFinder: Scientific Dataset Recommendation from Natural Language DescriptionsVijay Viswanathan, Luyu Gao, Tongshuang Wu, Pengfei Liu 等ACL 2023 · 被引用 9 次
相关 Paper
- AgentSelect: Benchmark for Narrative Query-to-Agent RecommendationYunxiao Shi, Wujiang Xu, Tingwei Chen, Haoning Shang 等ICML 2026 · 被引用 3 次
- Optimizing Retrieval for RAG via Reinforcement LearningJiawei Zhou, Lei ChenNeurIPS 2025 · 被引用 1 次
- From Reproduction to Replication: Evaluating Research Agents with Progressive Code MaskingGyeongwon James Kim, Alex Wilf, Louis-Philippe Morency, Daniel FriedICLR 2026 · 被引用 12 次
- AgentDR: Dynamic Recommendation with Implicit Item-Item Relations via LLM-based AgentsMingdai Yang, Nurendra Choudhary, Jiangshu Du, Edward W. Huang 等WWW 2026
- ReRec: Reasoning-Augmented LLM-based Recommendation Assistant via Reinforcement Fine-tuningJiani Huang, Shijie Wang, Liang-Bo Ning, Wenqi Fan 等ACL 2026 · 被引用 1 次
