ToolScope: Enhancing LLM Agent Tool Use through Tool Merging and Context-Aware Filtering
Marianne Menglin Liu, Daniel Garcia, Fjona Parllaku, Vikas Upadhyay, Fahad Shah, Dan Roth
摘要
Large language model (LLM) agents rely on external tools to solve complex tasks. However, real-world toolsets often contain semantically redundant tools with overlapping names and descriptions, introducing ambiguity and degrading tool selection performance. In addition, LLMs face strict input context limits, which prevent the agent from efficiently considering a large number of tools per query. To address these challenges, we propose ToolScope, a novel approach which contains: (1) ToolScope-Merger with Auto-Correction: automatically audits and fixes tool merges, reducing semantic redundancy in large toolsets. (2) ToolScop-eRetriever, which ranks and selects only the top-k relevant tools for a given query, effectively compressing the toolset to fit within the LLM's input window without sacrificing selection accuracy. This selective filtering directly mitigates context length constraints by ensuring that only the most relevant tools are passed to the model. We evaluate ToolScope using 3 state-of-the-art LLMs across 3 open-source tool-use benchmarks covering both single-tool and multi-tool scenarios in diverse real-world domains. Experimental results show a substantial increase of 8.38% to 38.6% in tool selection accuracy, demonstrating ToolScope's effectiveness in enhancing LLM tool-use capabilities.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Gorilla: Large Language Model Connected with Massive APIsShishir G. Patil, Tianjun Zhang, Xin Wang, Joseph E. GonzalezNeurIPS 2024 · 被引用 1,715 次
- ToolRL: Reward is All Tool Learning NeedsCheng Qian, Emre Can Acikgoz, Qi He, Hongru Wang 等NeurIPS 2025 · 被引用 387 次
- GPT4Tools: Teaching Large Language Model to Use Tools via Self-instructionRui Yang, Lin Song, Yanwei Li, Sijie Zhao 等NeurIPS 2023 · 被引用 340 次
- CRAFT: Customizing LLMs by Creating and Retrieving from Specialized ToolsetsLifan Yuan, Yangyi Chen, Xingyao Wang, Yi Fung 等ICLR 2024 · 被引用 117 次
- AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive ReasoningShirley Wu, Shiyu Zhao, Qian Huang, Kexin Huang 等NeurIPS 2024 · 被引用 95 次
相关 Paper
- BiasBusters: Uncovering and Mitigating Tool Selection Bias in Large Language ModelsThierry Blankenstein, Jialin Yu, Zixuan Li, Vassilis Plachouras 等ICLR 2026 · 被引用 8 次
- Tool Learning in the Wild: Empowering Language Models as Automatic Tool AgentsZhengliang Shi, Shen Gao, Lingyong Yan, Yue Feng 等WWW 2025 · 被引用 59 次
- Tools are under-documented: Simple Document Expansion Boosts Tool RetrievalXuan Lu, Haohang Huang, Rui Meng, Yaohui Jin 等ICLR 2026 · 被引用 16 次
- MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to UseYue Huang, Jiawen Shi, Yuan Li, Chenrui Fan 等ICLR 2024 · 被引用 188 次
- TRAJECT-Bench: A Trajectory-Aware Benchmark for Evaluating Agentic Tool UsePengfei He, Zhenwei Dai, Bing He, Hui Liu 等ICLR 2026 · 被引用 46 次
