SafeSieve: From Heuristics to Experience in Progressive Pruning for LLM-based Multi-Agent Communication
Ruijia Zhang, Xinyan Zhao, Ruixiang Wang, Sigen Chen, Guibin Zhang, An Zhang, Kun Wang, Qingsong Wen
Abstract
LLM-based multi-agent systems exhibit strong collaborative capabilities but often suffer from redundant communication and excessive token overhead. Existing methods typically enhance efficiency through pretrained GNNs or greedy algorithms, but often isolate pre-and post-task optimization, lacking a unified strategy. To this end, we present Safe-Sieve, a progressive and adaptive multi-agent pruning algorithm that dynamically refines the inter-agent communication through a novel dual-mechanism. SafeSieve integrates initial LLM-based semantic evaluation with accumulated performance feedback, enabling a smooth transition from heuristic initialization to experience-driven refinement. Unlike existing greedy Top-k pruning methods, SafeSieve employs 0-extension clustering to preserve structurally coherent agent groups while eliminating ineffective links. Experiments across benchmarks (SVAMP, HumanEval, etc.) showcase that SafeSieve achieves 94.01% average accuracy while reducing token usage by 12.4%-27.8%. Results further demonstrate robustness under prompt injection attacks (1.23% average accuracy drop). In heterogeneous settings, SafeSieve reduces deployment costs by 13.3% while maintaining performance. These results establish SafeSieve as an efficient, GPU-free, and scalable framework for practical multi-agent systems. Our code can be found below.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 210b5e6a-1252-4f2c-9ffe-6d2e9d4679b9Cited by top-tier papers3
- Stop Wasting Your Tokens: Towards Efficient Runtime Multi-Agent SystemsFulin Lin, Shaowen Chen, Ruishan Fang, Hongwei Wang et al.ICLR 2026 · 15 citations
- AgentSlimming: Towards Efficient and Cost-Aware Multi-Agent SystemsYulang Chen, Haoxuan Peng, Jinyan Liu, Zichen Wen et al.ACL 2026 · 1 citation
- Hetero-Designer: Automated Design of Multi-Agent Systems with Heterogeneous LLMsZhiheng Zhang, Yuanzhe Zhang, Bohan Yu, Daojian Zeng et al.ACL 2026
Builds on10
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou et al.ICLR 2021 · 7,905 citations
- Let's Verify Step by StepHunter Lightman, Vineet Kosaraju, Yuri Burda, Harrison Edwards et al.ICLR 2024 · 3,045 citations
- Chain of Agents: Large Language Models Collaborating on Long-Context TasksYusen Zhang, Ruoxi Sun, Yanfei Chen, Tomas Pfister et al.NeurIPS 2024 · 297 citations
- GPTSwarm: Language Agents as Optimizable GraphsMingchen Zhuge, Wenyi Wang, Louis Kirsch, Francesco Faccio et al.ICML 2024 · 45 citations
Related papers
- Cut the Crap: An Economical Communication Pipeline for LLM-based Multi-Agent SystemsGuibin Zhang, Yanwei Yue, Zhixun Li, Sukwon Yun et al.ICLR 2025
- G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural NetworksGuibin Zhang, Yanwei Yue, Xiangguo Sun, Guancheng Wan et al.ICML 2025
- AgentTailor: A Semantic-Aware LLM-Based Multi-Agent System with Actor-Critic StructurePeiting Yang, Jiahao Shi, Caiyi Xu, Ming Liu et al.ICML 2026
- Cost-Effective Communication: An Auction-based Method for Language Agent InteractionYijia Fan, Jusheng Zhang, Kaitong Cai, Jing Yang et al.AAAI 2026 · 15 citations
- Understanding and Mitigating Token-Pruning-Induced Vulnerabilities in VLMsShuailong Wang, Xinyu Lyu, Shengming Yuan, Jingkuan Song et al.ICML 2026
