Jigsaw: Supporting Designers to Prototype Multimodal Applications by Chaining AI Foundation Models
David Chuan-En Lin, Nikolas Martelaro
Abstract
Recent advancements in AI foundation models have made it possible for them to be utilized off-the-shelf for creative tasks, including ideating design concepts or generating visual prototypes. However, integrating these models into the creative process can be challenging as they often exist as standalone applications tailored to specific tasks. To address this challenge, we introduce Jigsaw, a prototype system that employs puzzle pieces as metaphors to represent foundation models. Jigsaw allows designers to combine different foundation model capabilities across various modalities by assembling compatible puzzle pieces. To inform the design of Jigsaw, we interviewed ten designers and distilled design goals. In a user study, we showed that Jigsaw enhanced designers’ understanding of available foundation model capabilities, provided guidance on combining capabilities across different modalities and tasks, and served as a canvas to support design exploration, prototyping, and documentation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature ReviewRock Yuren Pang, Hope Schroeder, Kynnedy Simone Smith, Solon Barocas et al.CHI 2025 · 51 citations
- How CO2STLY Is CHI? The Carbon Footprint of Generative AI in HCI Research and What We Should Do About ItNanna Inie, Jeanette Falk, Raghavendra SelvanCHI 2025 · 33 citations
- DesignWeaver: Dimensional Scaffolding for Text-to-Image Product DesignSirui Tao, Ivan Liang, Cindy Peng, Zhiqing Wang et al.CHI 2025 · 19 citations
- DynEx: Dynamic Code Synthesis with Structured Design Exploration for Accelerated Exploratory ProgrammingJenny Guangzhen Ma, Karthik Sreedhar, Vivian Liu, Pedro Alejandro Perez et al.CHI 2025 · 11 citations
- RemiHaven: Integrating "In-Town" and "Out-of-Town" Peers to Provide Personalized Reminiscence Support for Older DriftersXuechen Zhang, Changyang He, Peng Zhang, Hansu Gu et al.CHI 2025 · 5 citations
Builds on19
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
Related papers
- Creative Blends of Visual ConceptsZhida Sun, Zhenyao Zhang, Yue Zhang, Min Lu et al.CHI 2025 · 11 citations
- Relational Programming with Foundational ModelsZiyang Li, Jiani Huang, Jason Liu, Felix Zhu et al.AAAI 2024 · 11 citations
- ProductMeta: An Interactive System for Metaphorical Product Design Ideation with Multimodal Large Language ModelsQinyi Zhou, Jie Deng, Yu Liu, Yun Wang et al.CHI 2025 · 13 citations
- GenQuery: Supporting Expressive Visual Search with Generative ModelsKihoon Son, DaEun Choi, Tae Soo Kim, Young-Ho Kim et al.CHI 2024 · 48 citations
- FashionQ: An AI-Driven Creativity Support Tool for Facilitating Ideation in Fashion DesignYoungseung Jeon, Seungwan Jin, Patrick C. Shih, Kyungsik HanCHI 2021 · 143 citations
