EvAlignUX: Advancing UX Evaluation through LLM-Supported Metrics Exploration
Qingxiao Zheng, Minrui Chen, Pranav Sharma, Yiliu Tang, Mehul Oswal, Yiren Liu, Yun Huang
摘要
Evaluating UX in the context of AI’s complexity, unpredictability, and generative nature presents unique challenges. How can we support HCI researchers to create comprehensive UX evaluation plans? In this paper, we introduce EvAlignUX , a system powered by large language models and grounded in scientific literature, designed to help HCI researchers explore evaluation metrics and their relationship to research outcomes. A user study with 19 HCI scholars showed that EvAlignUX improved the perceived quality and confidence in UX evaluation plans while prompting deeper consideration of research impact and risks. The system enhanced participants’ thought processes, leading to the creation of a “UX Question Bank” to guide UX evaluation development. Findings also highlight how researchers’ backgrounds influence their inspiration and concerns about AI over-reliance, pointing to future research on AI’s role in fostering critical thinking. In a world where experience defines impact, we discuss the importance of shifting UX evaluation from a “method-centric” to a “mindset-centric” approach as the key to meaningful and lasting design evaluation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Reporting and Reviewing LLM-Integrated Systems in HCI: Challenges and ConsiderationsKarla Felix Navarro, Eugene Syriani, Ian ArawjoCHI 2026 · 被引用 1 次
- An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering SystemsAnna Martin-Boyle, William Humphreys, Martha Brown, Cara A. C. Leckey 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper17
- Questioning the AI: Informing Design Practices for Explainable AI User ExperiencesQ. Vera Liao, Daniel M. Gruen, Sarah MillerCHI 2020 · 被引用 758 次
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok 等CHI 2021 · 被引用 713 次
- Re-examining Whether, Why, and How Human-AI Interaction Is Uniquely Difficult to DesignQian Yang, Aaron Steinfeld, Carolyn P. Rosé, John ZimmermanCHI 2020 · 被引用 604 次
- Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human PreferencesShreya Shankar, J. D. Zamfirescu-Pereira, Bjoern Hartmann, Aditya G. Parameswaran 等UIST 2024 · 被引用 143 次
- UX Research on Conversational Human-AI Interaction: A Literature Review of the ACM Digital LibraryQingxiao Zheng, Yiliu Tang, Yiren Liu, Weizi Liu 等CHI 2022 · 被引用 99 次
相关 Paper
- How AI Processing Delays Foster Creativity: Exploring Research Question Co-Creation with an LLM-based AgentYiren Liu, Si Chen, Haocong Cheng, Mengxia Yu 等CHI 2024 · 被引用 57 次
- EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined CriteriaTae Soo Kim, Yoonjoo Lee, Jamin Shin, Young-Ho Kim 等CHI 2024 · 被引用 81 次
- Analyzing Collaborative Challenges and Needs of UX Practitioners when Designing with AI/MLMeena Devii Muralikumar, David W. McDonaldCSCW 2024 · 被引用 6 次
- HILL: A Hallucination Identifier for Large Language ModelsFlorian Leiser, Sven Eckhardt, Valentin Leuthe, Merlin Knaeble 等CHI 2024 · 被引用 67 次
- Systematic Task Exploration with LLMs: A Study in Citation Text GenerationFurkan Sahinuç, Ilia Kuznetsov, Yufang Hou, Iryna GurevychACL 2024
