CREA: A Collaborative Multi-Agent Framework for Creative Image Editing and Generation
Kavana Venkatesh, Connor Dunlop, Pinar Yanardag
摘要
Creativity in AI imagery remains a fundamental challenge, requiring not only the generation of visually compelling content but also the capacity to add novel, expressive, and artistically rich transformations to images. Unlike conventional editing tasks that rely on direct prompt-based modifications, creative image editing requires an autonomous, iterative approach that balances originality, coherence, and artistic intent. To address this, we introduce CREA, a novel multi-agent collaborative framework that mimics the human creative process. Our framework leverages a team of specialized AI agents who dynamically collaborate to conceptualize, generate, critique, and enhance images. Through extensive qualitative and quantitative evaluations, we demonstrate that CREA significantly outperforms state-of-the-art methods in diversity, semantic alignment, and creative transformation. To the best of our knowledge, this is the first work to introduce the task of creative editing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- CREward: A Type-Specific Creativity Reward ModelJiyeon Han, Ali Mahdavi-Amiri, Hao Zhang, Haedong JeongCVPR 2026
- OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided FuzzingJianming Chen, Yawen Wang, Junjie Wang, Zhe Liu 等ICML 2026
它引用的顶会 Paper23
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- ContextCam: Bridging Context Awareness with Creative Human-AI Image Co-CreationXianzhe Fan, Zihan Wu, Chun Yu, Fenggui Rao 等CHI 2024 · 被引用 49 次
- Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic GuidanceHongxing Fan, Lipeng Wang, Haohua Chen, Zehuan Huang 等ACM MM 2025 · 被引用 3 次
- Creation-Mmbench: Assessing Context-Aware Creative Intelligence in MllmsXinyu Fang, Zhijian Chen, Kai Lan, Lixin Ma 等ICCV 2025 · 被引用 23 次
- Talk2Image: A Multi-Agent System for Multi-Turn Image Generation and EditingShichao Ma, Yunhe Guo, Jiahao Su, Qihe Huang 等AAAI 2026 · 被引用 9 次
- Understanding Nonlinear Collaboration between Human and AI Agents: A Co-design Framework for Creative DesignJiayi Zhou, Renzhong Li, Junxiu Tang, Tan Tang 等CHI 2024 · 被引用 100 次
