Heterogeneous Swarms: Jointly Optimizing Model Roles and Weights for Multi-LLM Systems
Shangbin Feng, Zifeng Wang, Palash Goyal, Yike Wang, Weijia Shi, Huang Xia, Hamid Palangi, Luke Zettlemoyer, Yulia Tsvetkov, Chen-Yu Lee, Tomas Pfister
摘要
We propose HETEROGENEOUS SWARMS, an algorithm to design multi-LLM systems by jointly optimizing model roles and weights. We represent multi-LLM systems as directed acyclic graphs (DAGs) of LLMs with topological message passing for collaborative generation. Given a pool of LLM experts and a utility function, HETEROGENEOUS SWARMS employs two iterative steps: role-step and weight-step. For role-step, we interpret model roles as learning a DAG that specifies the flow of inputs and outputs between LLMs. Starting from a swarm of random continuous adjacency matrices, we decode them into discrete DAGs, call the LLMs in topological order, evaluate on the utility function (e.g. accuracy on a task), and optimize the adjacency matrices with particle swarm optimization based on the utility score. For weight-step, we assess the contribution of individual LLMs in the multi-LLM systems and optimize model weights with swarm intelligence. We propose JFK-score to quantify the individual contribution of each LLM in the best-found DAG of the role-step, then optimize model weights with particle swarm optimization based on the JFK-score. Experiments demonstrate that HETEROGE-NEOUS SWARMS outperforms 17 role-and/or weight-based baselines by 18.5% on average across 12 tasks. Further analysis reveals that HETEROGENEOUS SWARMS discovers multi-LLM systems with heterogeneous model roles and substantial collaborative gains, and benefits from the diversity of language models. 2 * Work done as a student researcher at Google Cloud AI Research.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- FlexOLMo: Open Language Models for Flexible Data UseWeijia Shi, Akshita Bhagia, Kevin Farhat, Niklas Muennighoff 等NeurIPS 2025 · 被引用 16 次
- Learning Decentralized LLM Collaboration with Multi-Agent Actor CriticShuo Liu, Tianle Chen, Ryan Amiri, Christopher AmatoICML 2026 · 被引用 6 次
- S-DAG: A Subject-Based Directed Acyclic Graph for Multi-Agent Heterogeneous ReasoningJiangwen Dong, Zehui Lin, Wanyu Lin, Mingjin ZhangAAAI 2026 · 被引用 4 次
- VeriTrail: Closed-Domain Hallucination Detection with TraceabilityDasha Metropolitansky, Jonathan LarsonICLR 2026 · 被引用 3 次
- Among Us: Measuring and Mitigating Malicious Contributions in Model Collaboration SystemsZiyuan Yang, Wenxuan Ding, Shangbin Feng, Yulia TsvetkovACL 2026 · 被引用 1 次
它引用的顶会 Paper80
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes 等ICLR 2020 · 被引用 4,112 次
- Improving Factuality and Reasoning in Language Models through Multiagent DebateYilun Du, Shuang Li, Antonio Torralba, Joshua B. Tenenbaum 等ICML 2024 · 被引用 1,562 次
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
相关 Paper
- Model Swarms: Collaborative Search to Adapt LLM Experts via Swarm IntelligenceShangbin Feng, Zifeng Wang, Yike Wang, Sayna Ebrahimi 等ICML 2025
- Hetero-Designer: Automated Design of Multi-Agent Systems with Heterogeneous LLMsZhiheng Zhang, Yuanzhe Zhang, Bohan Yu, Daojian Zeng 等ACL 2026
- GPTSwarm: Language Agents as Optimizable GraphsMingchen Zhuge, Wenyi Wang, Louis Kirsch, Francesco Faccio 等ICML 2024 · 被引用 45 次
- HieraMAS: Optimizing Intra-Node LLM Mixtures and Inter-Node Topology for Multi-Agent SystemsTianjun Yao, Zhaoyi Li, Zhiqiang ShenICML 2026 · 被引用 1 次
- Advancing Collaborative Debates with Role Differentiation through Multi-Agent Reinforcement LearningHaoran Li, Ziyi Su, Yun Xue, Zhiliang Tian 等ACL 2025 · 被引用 8 次
