Alignment for Efficient Tool Calling of Large Language Models
Hongshen Xu, Zihan Wang, Zichen Zhu, Lei Pan, Xingyu Chen, Shuai Fan, Lu Chen, Kai Yu
Abstract
Recent advancements in tool learning have enabled large language models (LLMs) to integrate external tools, enhancing their task performance by expanding their knowledge boundaries. However, relying on tools often introduces trade-offs between performance, speed, and cost, with LLMs sometimes exhibiting overreliance and overconfidence in tool usage. This paper addresses the challenge of aligning LLMs with their knowledge boundaries to make more intelligent decisions about tool invocation. We propose a multi-objective alignment framework that combines probabilistic knowledge boundary estimation with dynamic decision-making, allowing LLMs to better assess when to invoke tools based on their confidence. Our framework includes two methods for knowledge boundary estimation-consistency-based and absolute estimation-and two training strategies for integrating these estimates into the model's decision-making process. Experimental results on various tool invocation scenarios demonstrate the effectiveness of our framework, showing significant improvements in tool efficiency by reducing unnecessary tool usage.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a6dbf2ed-a663-4db7-af60-bf44f50ea1c8Cited by top-tier papers5
- DiSRouter: Distributed Self-Routing for LLM SelectionsHang Zheng, Hongshen Xu, Yongkai.lin, Shuai Fan et al.ICLR 2026 · 6 citations
- Learning from Contrasts: Synthesizing Reasoning Paths from Diverse Search TrajectoriesPeiyang Liu, Zhirui Chen, Xi Wang, Di Liang et al.ACL 2026 · 4 citations
- DiffoR: A Unified Continuous Generative Framework for Universal Ordinal RegressionHongxu Ma, Lin Wang, Chenghou Jin, Han Zhou et al.KDD 2026 · 1 citation
- Reducing Tool Hallucination via Reliability AlignmentHongshen Xu, Zichen Zhu, Lei Pan, Zihan Wang et al.ICML 2025
- DORA: A Dual-Objective Reinforcement Learning Framework for Effective and Efficient Multimodal Agentic SearchGuangming Qin, Yuhao Deng, Yukun Zhao, Zhenyang Li et al.ACL 2026
Builds on16
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Toolformer: Language Models Can Teach Themselves to Use ToolsTimo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu et al.NeurIPS 2023 · 5,989 citations
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu et al.ICLR 2022 · 4,966 citations
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach et al.ICLR 2022 · 1,976 citations
Related papers
- Adaptive Tool Use in Large Language Models with Meta-Cognition TriggerWenjun Li, Dexun Li, Kuicai Dong, Cong Zhang et al.ACL 2025
- Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary PerceptionShiyu Ni, Keping Bi, Jiafeng Guo, Lulu Yu et al.ACL 2025 · 22 citations
- Tool Learning in the Wild: Empowering Language Models as Automatic Tool AgentsZhengliang Shi, Shen Gao, Lingyong Yan, Yue Feng et al.WWW 2025 · 59 citations
- Towards Pareto-Optimal Tool-Integrated Agents with Pareto Ranking Policy OptimizationJunyi Li, Xiaowei Qian, Yingyi Zhang, Wenlin Zhang et al.ICML 2026
- MASH: Modeling Abstention via Selective Help-SeekingMustafa Omer Gul, Claire Cardie, Tanya GoyalICML 2026 · 2 citations
