Teaching Humans When to Defer to a Classifier via Exemplars
Hussein Mozannar, Arvind Satyanarayan, David A. Sontag
摘要
Expert decision makers are starting to rely on data-driven automated agents to assist them with various tasks. For this collaboration to perform properly, the human decision maker must have a mental model of when and when not to rely on the agent. In this work, we aim to ensure that human decision makers learn a valid mental model of the agent's strengths and weaknesses. To accomplish this goal, we propose an exemplar-based teaching strategy where humans solve a set of selected examples and with our help generalize from them to the domain. We present a novel parameterization of the human's mental model of the AI that applies a nearest neighbor rule in local regions surrounding the teaching examples. Using this model, we derive a near-optimal strategy for selecting a representative teaching set. We validate the benefits of our teaching strategy on a multi-hop question answering task with an interpretable AI model using crowd workers. We find that when workers draw the right lessons from the teaching stage, their task performance improves. We furthermore validate our method on a set of synthetic experiments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-MakingShuai Ma, Ying Lei, Xinru Wang, Chengbo Zheng 等CHI 2023 · 被引用 139 次
- Improving Human-AI Partnerships in Child Welfare: Understanding Worker Practices, Challenges, and Desires for Algorithmic Decision SupportAnna Kawakami, Venkatesh Sivaraman, Hao Fei Cheng, Logan Stapleton 等CHI 2022 · 被引用 137 次
- Exploring Challenges and Opportunities to Support Designers in Learning to Co-create with AI-based Manufacturing Design ToolsFrederic Gmeiner, Humphrey Yang, Lining Yao, Kenneth Holstein 等CHI 2023 · 被引用 125 次
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 被引用 98 次
- Improving Human-AI Collaboration With Descriptions of AI BehaviorÁngel Alexander Cabrera, Adam Perer, Jason I. HongCSCW 2023 · 被引用 85 次
它引用的顶会 Paper8
- A Human-Centered Evaluation of a Deep Learning System Deployed in Clinics for the Detection of Diabetic RetinopathyEmma Beede, Elizabeth Elliott Baylor, Fred Hersch, Anna Iurchenko 等CHI 2020 · 被引用 589 次
- Interpreting Interpretability: Understanding Data Scientists' Use of Interpretability Tools for Machine LearningHarmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana 等CHI 2020 · 被引用 541 次
- Evaluating Explainable AI: Which Algorithmic Explanations Help Users Predict Model Behavior?Peter Hase, Mohit BansalACL 2020 · 被引用 216 次
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz 等AAAI 2021 · 被引用 185 次
- Select, Answer and Explain: Interpretable Multi-Hop Reading Comprehension over Multiple DocumentsMing Tu, Kevin Huang, Guangtao Wang, Jing Huang 等AAAI 2020 · 被引用 155 次
相关 Paper
- Learning to Explain Selectively: A Case Study on Question AnsweringShi Feng, Jordan L. Boyd-GraberEMNLP 2022 · 被引用 4 次
- Effective Human-AI Teams via Learned Natural Language Rules and OnboardingHussein Mozannar, Jimin J. Lee, Dennis Wei, Prasanna Sattigeri 等NeurIPS 2023 · 被引用 28 次
- 'The AI is uncertain, so am I. What now?': Navigating Shortcomings of Uncertainty Representations in Human-AI Collaboration with Capability-focused GuidanceUlrike Schäfer, Lars Sipos, Claudia Müller-BirnCSCW 2025 · 被引用 6 次
- Teachable Conversational Agents for Crowdwork: Effects on Performance and TrustNalin Chhibber, Joslin Goh, Edith LawCSCW 2022 · 被引用 2 次
- Crowd Teaching with Imperfect LabelsYao Zhou, Arun Reddy Nelakurthi, Ross Maciejewski, Wei Fan 等WWW 2020 · 被引用 12 次
