Collaposer: Transforming Photo Collections into Visual Assets for Storytelling with Collages
Jiayi Zhou, Liwenhan Xie, Jiaju Ma, Zheng Wei, Huamin Qu, Anyi Rao
摘要
Digital collage is an artistic practice that combines image cutouts to tell stories. However, preparing cutouts from a set of photos remains a tedious and time-consuming task. A formative study identified three main challenges: 1) inefficient search for relevant photos, 2) manual image cutout, and 3) difficulty in organizing large sets of cutouts. To meet these challenges and facilitate asset preparation for collage, we propose Collaposer, a tool that transforms a collection of photos into organized, ready-to-use visual cutouts based on user-provided story descriptions. Collaposer tags, detects, and segments photos, and then uses an LLM to select central and related labels based on the user-provided story description. Collaposer presents the resulting visuals in varying sizes, clustered according to semantic hierarchy. Our evaluation shows that Collaposer effectively automates the preparation process to produce diverse sets of visual cutouts adhering to the storyline, allowing users to focus on collaging these assets for storytelling.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Bridging the Gulf of Envisioning: Cognitive Challenges in Prompt Based Interactions with LLMsHariharan Subramonyam, Roy Pea, Christopher Lawrence Pondoc, Maneesh Agrawala 等CHI 2024 · 被引用 137 次
- Understanding Nonlinear Collaboration between Human and AI Agents: A Co-design Framework for Creative DesignJiayi Zhou, Renzhong Li, Junxiu Tang, Tan Tang 等CHI 2024 · 被引用 100 次
- VINS: Visual Search for Mobile User Interface DesignSara Bunian, Kai Li, Chaima Jemmali, Casper Harteveld 等CHI 2021 · 被引用 100 次
- MetaMap: Supporting Visual Metaphor Ideation through Multi-dimensional Example-based ExplorationYouwen Kang, Zhida Sun, Sitong Wang, Zeyu Huang 等CHI 2021 · 被引用 78 次
相关 Paper
- CollageNoter: Real-Time and Adaptive Collage Layout Design for Screenshot-Based E-Note-TakingQiuyun Zhang, Bin Guo, Lina Yao, Xiaotian Qiao 等AAAI 2025
- SCORE: Semantic Collage by Optimizing Rendered ElementsZefan Shao, Jin Zhou, Hongliang Yang, Pengfei XuAAAI 2026
- PhotoScout: Synthesis-Powered Multi-Modal Image SearchCeleste Barnaby, Qiaochu Chen, Chenglong Wang, Isil DilligCHI 2024 · 被引用 9 次
- MapStory: Prototyping Editable Map Animations with LLM AgentsAditya Gunturu, Ben Pearman, Keiichi Ihara, Morteza Faraji 等UIST 2025
- Opal: Multimodal Image Generation for News IllustrationVivian Liu, Han Qiao, Lydia B. ChiltonUIST 2022 · 被引用 97 次
