Creativity in LLM-based Multi-Agent Systems: A Survey
Yi-Cheng Lin, Kang-Chieh Chen, Zhe-Yan Li, Tzu-Heng Wu, Tzu-Hsuan Wu, Kuan-Yu Chen, Hung-yi Lee, Yun-Nung Chen
Abstract
Large language model (LLM)-driven multiagent systems (MAS) are transforming how humans and AIs collaboratively generate ideas and artifacts. While existing surveys provide comprehensive overviews of MAS infrastructures, they largely overlook the dimension of creativity, including how novel outputs are generated and evaluated, how creativity informs agent personas, and how creative workflows are coordinated. This is the first survey dedicated to creativity in MAS. We focus on text and image generation tasks, and present: (1) a taxonomy of agent proactivity and persona design; (2) an overview of generation techniques, including divergent exploration, iterative refinement, and collaborative synthesis, as well as relevant datasets and evaluation metrics; and (3) a discussion of key challenges, such as inconsistent evaluation standards, insufficient bias mitigation, coordination conflicts, and the lack of unified benchmarks. This survey offers a structured framework and roadmap for advancing the development, evaluation, and standardization of creative MAS. 1 * These authors contributed equally. 1 https://github.com/MiuLab/MultiAgent-Survey You are Jack, a 35-year-old male creative technologist and research fellow at the Institute for Human-AI Creative Synergy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 14bc3fa4-b86e-4d23-9fca-07a8f7b4994dCited by top-tier papers5
- Understanding Human-Multi-Agent Team Formation for Creative WorkHyunseung Lim, Dasom Choi, Sooyohn Nam, Bogoan Kim et al.CHI 2026 · 2 citations
- TraceCoder: A Trace-Driven Multi-Agent Framework for Automated Debugging of LLM-Generated CodeJiangping Huang, Wenguang Ye, Weisong Sun, Jian Zhang et al.ICSE 2026 · 1 citation
- WRitEer: A Multi-Objective, Preference-Driven Multi-Agent Framework for Human-Like Advanced Text GenerationJunchuan Yu, Yuyang SunAAAI 2026
- TRACE: A Corpus of Team Creative DiscussionsYixuan Jiang, Tiancheng Hu, José Hernández-Orallo, David Stillwell et al.ACL 2026
- When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent SystemsZehao Wang, shilong jin, Zhao Cao, Lanjun WangICML 2026
Builds on29
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris et al.UIST 2023 · 1,882 citations
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human PreferenceWei-Lin Chiang, Lianmin Zheng, Ying Sheng, Anastasios Nikolas Angelopoulos et al.ICML 2024 · 1,212 citations
- Bias Runs Deep: Implicit Reasoning Biases in Persona-Assigned LLMsShashank Gupta, Vaishnavi Shrivastava, Ameet Deshpande, Ashwin Kalyan et al.ICLR 2024 · 212 citations
Related papers
- Automated Creativity Evaluation of Language Models Across Open-Ended TasksTan Min Sen, Zachary Choy Kit Chun, Syed Ali Redha Alsagoff, Nadya Yuki Wangsajaya et al.ACL 2026
- Human Creativity in the Age of LLMs: Randomized Experiments on Divergent and Convergent ThinkingHarsh Kumar, Jonathan Vincentius, Ewan Jordan, Ashton AndersonCHI 2025 · 107 citations
- A Survey of Large Language Models for Text-Guided Molecular Discovery: From Molecule Generation to OptimizationZiqing Wang, Kexin Zhang, Zihan Zhao, Yibo Wen et al.ACL 2026 · 10 citations
- AI-Augmented Brainwriting: Investigating the use of LLMs in group ideationOrit Shaer, Angelora Cooper, Osnat Mokryn, Andrew L. Kun et al.CHI 2024 · 120 citations
- Systematic Task Exploration with LLMs: A Study in Citation Text GenerationFurkan Sahinuç, Ilia Kuznetsov, Yufang Hou, Iryna GurevychACL 2024
