Brickify: Enabling Expressive Design Intent Specification through Direct Manipulation on Design Tokens
Xinyu Shi, Yinghou Wang, Ryan A. Rossi, Jian Zhao
摘要
Expressing design intent using natural language prompts requires designers to verbalize the ambiguous visual details concisely, which can be challenging or even impossible. To address this, we introduce Brickify, a visual-centric interaction paradigm — expressing design intent through direct manipulation on design tokens. Brickify extracts visual elements (e.g., subject, style, and color) from reference images and converts them into interactive and reusable design tokens that can be directly manipulated (e.g., resize, group, link, etc.) to form the visual lexicon. The lexicon reflects users’ intent for both what visual elements are desired and how to construct them into a whole. We developed Brickify to demonstrate how AI models can interpret and execute the visual lexicon through an end-to-end pipeline. In a user study, experienced designers found Brickify more efficient and intuitive than text-based prompts, allowing them to describe visual details, explore alternatives, and refine complex designs with greater ease and control.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- StoryEnsemble: Enabling Dynamic Exploration & Iteration in the Design Process with AI and Forward-Backward PropagationSangho Suh, Michael Lai, Kevin Pu, Steven P. Dow 等UIST 2025 · 被引用 5 次
- DataWink: Reusing and Adapting SVG-Based Visualization Examples with Large Multimodal ModelsLiwenhan Xie, Yanna Lin, Can Liu, Huamin Qu 等IEEE VIS 2025 · 被引用 3 次
- DesignTrace: Exploring, Iterating and Tracking Design Alternatives with GenAIXiaohan Peng, Debanjana Haldar, Wendy E. Mackay, Janin KochCHI 2026 · 被引用 3 次
- Collaposer: Transforming Photo Collections into Visual Assets for Storytelling with CollagesJiayi Zhou, Liwenhan Xie, Jiaju Ma, Zheng Wei 等CHI 2026 · 被引用 3 次
- Vistoria: A Multimodal System to Support Fictional Story Writing through Instrumental Image-Text Co-EditingKexue Fu, Jingfei Huang, Long Ling, Sumin Hong 等CHI 2026 · 被引用 3 次
它引用的顶会 Paper44
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language ModelsStephen Brade, Bryan Wang, Maurício Sousa, Sageev Oore 等UIST 2023 · 被引用 179 次
- Bridging Gulfs in UI Generation through Semantic GuidanceSeokhyeon Park, Soohyun Lee, Eugene Choi, Hyunwoo Kim 等CHI 2026 · 被引用 3 次
- WorldSmith: Iterative and Expressive Prompting for World Building with a Generative AIHai Dang, Frederik Brudy, George W. Fitzmaurice, Fraser AndersonUIST 2023 · 被引用 43 次
- Interaction-Augmented Instruction: Modeling the Synergy of Prompts and Interactions in Human-GenAI CollaborationLeixian Shen, Yifang Wang, Huamin Qu, Xing Xie 等CHI 2026 · 被引用 3 次
- Selectively Extracting and Injecting Visual Attributes into Text-to-Image ModelsSeunghwan Choi, Jooyeol Yun, Youngdo Lee, Jaegul ChooCVPR 2026
