Planning for Natural Language Failures with the AI Playbook
Matthew K. Hong, Adam Fourney, Derek DeBellis, Saleema Amershi
摘要
Prototyping AI user experiences is challenging due in part to probabilistic AI models making it difficult to anticipate, test, and mitigate AI failures before deployment. In this work, we set out to support practitioners with early AI prototyping, with a focus on natural language (NL)-based technologies. Our interviews with 12 NL practitioners from a large technology company revealed that, in addition to challenges prototyping AI, prototyping was often not happening at all or focused only on idealized scenarios due to a lack of tools and tight timelines. These findings informed our design of the AI Playbook, an interactive and low-cost tool we developed to encourage proactive and systematic consideration of AI errors before deployment. Our evaluation of the AI Playbook demonstrates its potential to 1) encourage product teams to prioritize both ideal and failure scenarios, 2) standardize the articulation of AI failures from a user experience perspective, and 3) act as a boundary object between user experience designers, data scientists, and engineers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Investigating How Practitioners Use Human-AI Guidelines: A Case Study on the People + AI GuidebookNur Yildirim, Mahima Pushkarna, Nitesh Goyal, Martin Wattenberg 等CHI 2023 · 被引用 103 次
- Designing Responsible AI: Adaptations of UX Practice to Meet Responsible AI ChallengesQiaosi Wang, Michael Madaio, Shaun K. Kane, Shivani Kapania 等CHI 2023 · 被引用 90 次
- Designerly Understanding: Information Needs for Model Transparency to Support Design Ideation for AI-Powered User ExperienceQ. Vera Liao, Hariharan Subramonyam, Jennifer Wang, Jennifer Wortman VaughanCHI 2023 · 被引用 81 次
- Farsight: Fostering Responsible AI Awareness During AI Application PrototypingZijie J. Wang, Chinmay Kulkarni, Lauren Wilcox, Michael Terry 等CHI 2024 · 被引用 55 次
- Seamful XAI: Operationalizing Seamful Design in Explainable AIUpol Ehsan, Q. Vera Liao, Samir Passi, Mark O. Riedl 等CSCW 2024 · 被引用 41 次
它引用的顶会 Paper2
- Re-examining Whether, Why, and How Human-AI Interaction Is Uniquely Difficult to DesignQian Yang, Aaron Steinfeld, Carolyn P. Rosé, John ZimmermanCHI 2020 · 被引用 604 次
- Co-Designing Checklists to Understand Organizational Challenges and Opportunities around Fairness in AIMichael A. Madaio, Luke Stark, Jennifer Wortman Vaughan, Hanna M. WallachCHI 2020 · 被引用 428 次
相关 Paper
- fAIlureNotes: Supporting Designers in Understanding the Limits of AI Models for Computer Vision TasksSteven Moore, Q. Vera Liao, Hariharan SubramonyamCHI 2023 · 被引用 36 次
- Prototyping with Prompts: Emerging Approaches and Challenges in Generative AI Design for Collaborative Software TeamsHari Subramonyam, Divy Thakkar, Andrew Ku, Jürgen Dieber 等CHI 2025 · 被引用 26 次
- Zeno: An Interactive Framework for Behavioral Evaluation of Machine LearningÁngel Alexander Cabrera, Erica Fu, Donald Bertucci, Kenneth Holstein 等CHI 2023 · 被引用 51 次
- A Scoping Study of Evaluation Practices for Responsible AI Tools: Steps Towards Effectiveness EvaluationsGlen Berman, Nitesh Goyal, Michael MadaioCHI 2024 · 被引用 40 次
- Tinker, Tailor, Configure, Customize: The Articulation Work of Contextualizing an AI Fairness ChecklistMichael A. Madaio, Jingya Chen, Hanna M. Wallach, Jennifer Wortman VaughanCSCW 2024 · 被引用 13 次
