Collaposer: Transforming Photo Collections into Visual Assets for Storytelling with Collages
Jiayi Zhou, Liwenhan Xie, Jiaju Ma, Zheng Wei, Huamin Qu, Anyi Rao
Abstract
Digital collage is an artistic practice that combines image cutouts to tell stories. However, preparing cutouts from a set of photos remains a tedious and time-consuming task. A formative study identified three main challenges: 1) inefficient search for relevant photos, 2) manual image cutout, and 3) difficulty in organizing large sets of cutouts. To meet these challenges and facilitate asset preparation for collage, we propose Collaposer, a tool that transforms a collection of photos into organized, ready-to-use visual cutouts based on user-provided story descriptions. Collaposer tags, detects, and segments photos, and then uses an LLM to select central and related labels based on the user-provided story description. Collaposer presents the resulting visuals in varying sizes, clustered according to semantic hierarchy. Our evaluation shows that Collaposer effectively automates the preparation process to produce diverse sets of visual cutouts adhering to the storyline, allowing users to focus on collaging these assets for storytelling.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on14
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Bridging the Gulf of Envisioning: Cognitive Challenges in Prompt Based Interactions with LLMsHariharan Subramonyam, Roy Pea, Christopher Lawrence Pondoc, Maneesh Agrawala et al.CHI 2024 · 137 citations
- Understanding Nonlinear Collaboration between Human and AI Agents: A Co-design Framework for Creative DesignJiayi Zhou, Renzhong Li, Junxiu Tang, Tan Tang et al.CHI 2024 · 100 citations
- VINS: Visual Search for Mobile User Interface DesignSara Bunian, Kai Li, Chaima Jemmali, Casper Harteveld et al.CHI 2021 · 100 citations
- MetaMap: Supporting Visual Metaphor Ideation through Multi-dimensional Example-based ExplorationYouwen Kang, Zhida Sun, Sitong Wang, Zeyu Huang et al.CHI 2021 · 78 citations
Related papers
- CollageNoter: Real-Time and Adaptive Collage Layout Design for Screenshot-Based E-Note-TakingQiuyun Zhang, Bin Guo, Lina Yao, Xiaotian Qiao et al.AAAI 2025
- SCORE: Semantic Collage by Optimizing Rendered ElementsZefan Shao, Jin Zhou, Hongliang Yang, Pengfei XuAAAI 2026
- PhotoScout: Synthesis-Powered Multi-Modal Image SearchCeleste Barnaby, Qiaochu Chen, Chenglong Wang, Isil DilligCHI 2024 · 9 citations
- MapStory: Prototyping Editable Map Animations with LLM AgentsAditya Gunturu, Ben Pearman, Keiichi Ihara, Morteza Faraji et al.UIST 2025
- Opal: Multimodal Image Generation for News IllustrationVivian Liu, Han Qiao, Lydia B. ChiltonUIST 2022 · 97 citations
