Generating Automatic Feedback on UI Mockups with Large Language Models
Peitong Duan, Jeremy Warner, Yang Li, Bjoern Hartmann
摘要
Feedback on user interface (UI) mockups is crucial in design. However, human feedback is not always readily available. We explore the potential of using large language models for automatic feedback. Specifically, we focus on applying GPT-4 to automate heuristic evaluation, which currently entails a human expert assessing a UI’s compliance with a set of design guidelines. We implemented a Figma plugin that takes in a UI design and a set of written heuristics, and renders automatically-generated feedback as constructive suggestions. We assessed performance on 51 UIs using three sets of guidelines, compared GPT-4-generated design suggestions with those from human experts, and conducted a study with 12 expert designers to understand fit with existing practice. We found that GPT-4-based feedback is useful for catching subtle errors, improving text, and considering UI semantics, but feedback also decreased in utility over iterations. Participants described several uses for this plugin despite its imperfect suggestions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature ReviewRock Yuren Pang, Hope Schroeder, Kynnedy Simone Smith, Solon Barocas 等CHI 2025 · 被引用 51 次
- How CO2STLY Is CHI? The Carbon Footprint of Generative AI in HCI Research and What We Should Do About ItNanna Inie, Jeanette Falk, Raghavendra SelvanCHI 2025 · 被引用 33 次
- UIClip: A Data-driven Model for Assessing User Interface DesignJason Wu, Yi-Hao Peng, Xin Yue Amanda Li, Amanda Swearngin 等UIST 2024 · 被引用 29 次
- LogoMotion: Visually-Grounded Code Synthesis for Creating and Editing AnimationVivian Liu, Rubaiat Habib Kazi, Li-Yi Wei, Matthew Fisher 等CHI 2025 · 被引用 23 次
- UICrit: Enhancing Automated Design Evaluation with a UI Critique DatasetPeitong Duan, Chin-Yi Cheng, Gang Li, Bjoern Hartmann 等UIST 2024 · 被引用 22 次
它引用的顶会 Paper16
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan 等NeurIPS 2023 · 被引用 5,828 次
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach 等ICLR 2022 · 被引用 1,976 次
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris 等UIST 2023 · 被引用 1,882 次
相关 Paper
- Closing the Loop between User Stories and GUI Prototypes: An LLM-Based Assistant for Cross-Functional Integration in Software DevelopmentFelix Kretzer, Kristian Kolthoff, Christian Bartelt, Simone Paolo Ponzetto 等CHI 2025 · 被引用 18 次
- DesignRepair: Dual-Stream Design Guideline-Aware Frontend Repair with Large Language ModelsMingyue Yuan, Jieshan Chen, Zhenchang Xing, Aaron Quigley 等ICSE 2025 · 被引用 2 次
- Using an LLM to Help With Code UnderstandingDaye Nam, Andrew Macvean, Vincent J. Hellendoorn, Bogdan Vasilescu 等ICSE 2024 · 被引用 264 次
- "Create a Fear of Missing Out" - ChatGPT Implements Unsolicited Deceptive Designs in Generated Websites Without WarningVeronika Krauß, Mark McGill, Thomas Kosch, Yolanda Maira Thiel 等CHI 2025 · 被引用 17 次
- SimUser: Generating Usability Feedback by Simulating Various Users Interacting with Mobile ApplicationsWei Xiang, Hanfei Zhu, Suqi Lou, Xinli Chen 等CHI 2024 · 被引用 49 次
