Explanations, Fairness, and Appropriate Reliance in Human-AI Decision-Making
Jakob Schoeffer, Maria De-Arteaga, Niklas Kühl
摘要
In this work, we study the effects of feature-based explanations on distributive fairness of AI-assisted decisions, specifically focusing on the task of predicting occupations from short textual bios. We also investigate how any effects are mediated by humans’ fairness perceptions and their reliance on AI recommendations. Our findings show that explanations influence fairness perceptions, which, in turn, relate to humans’ tendency to adhere to AI recommendations. However, we see that such explanations do not enable humans to discern correct and incorrect AI recommendations. Instead, we show that they may affect reliance irrespective of the correctness of AI recommendations. Depending on which features an explanation highlights, this can foster or hinder distributive fairness: when explanations highlight features that are task-irrelevant and evidently associated with the sensitive attribute, this prompts overrides that counter AI recommendations that align with gender stereotypes. Meanwhile, if explanations appear task-relevant, this induces reliance behavior that reinforces stereotype-aligned errors. These results imply that feature-based explanations are not a reliable mechanism to improve distributive fairness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language ModelsJessica Y. Bo, Sophia Wan, Ashton AndersonCHI 2025 · 被引用 31 次
- From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered AnalysisZhuoyan Li, Hangxiao Zhu, Zhuoran Lu, Ziang Xiao 等CHI 2025 · 被引用 30 次
- User Experience with LLM-powered Conversational Recommendation Systems: A Case of Music RecommendationSojeong Yun, Youn-kyung LimCHI 2025 · 被引用 13 次
- Understanding the Effects of AI-based Credibility Indicators When People Are Influenced By Both Peers and ExpertsZhuoran Lu, Patrick Li, Weilong Wang, Ming YinCHI 2025 · 被引用 6 次
- Do People Appropriately Rely on AI-Advice? An Analytical Review of HCI Research on Human-AI Decision-MakingMuhammad Raees, Vassilis-Javed Khan, Ioanna Lykourentzou, Konstantinos PapangelisCHI 2026 · 被引用 6 次
它引用的顶会 Paper16
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-makingZana Buçinca, Maja Barbara Malaya, Krzysztof Z. GajosCSCW 2021 · 被引用 962 次
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok 等CHI 2021 · 被引用 713 次
- Manipulating and Measuring Model InterpretabilityForough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman, Jennifer Wortman Vaughan 等CHI 2021 · 被引用 663 次
- Explanations Can Reduce Overreliance on AI Systems During Decision-MakingHelena Vasconcelos, Matthew Jörke, Madeleine Grunde-McLaughlin, Tobias Gerstenberg 等CSCW 2023 · 被引用 362 次
- How to Evaluate Trust in AI-Assisted Decision Making? A Survey of Empirical MethodologiesOleksandra Vereschak, Gilles Bailly, Baptiste CaramiauxCSCW 2021 · 被引用 227 次
相关 Paper
- The Explanation That Hits Home: The Characteristics of Verbal Explanations That Affect Human Perception in Subjective Decision-MakingSharon A. Ferguson, Paula Akemi Aoyagui, Rimsha Rizvi, Young-Ho Kim 等CSCW 2024 · 被引用 14 次
- Effect of Information Presentation on Fairness Perceptions of Machine Learning PredictorsNiels van Berkel, Jorge Gonçalves, Daniel Russo, Simo Hosio 等CHI 2021 · 被引用 84 次
- Mitigating Gender Stereotypes Toward AI Agents Through an eXplainable AI (XAI) ApproachWen Duan, Nathan J. McNeese, Guo Freeman, Lingyuan LiCSCW 2024 · 被引用 16 次
- On the Mutual Influence of Gender and Occupation in LLM RepresentationsHaozhe An, Connor Baumler, Abhilasha Sancheti, Rachel RudingerACL 2025
- Towards Conceptualization of "Fair Explanation": Disparate Impacts of anti-Asian Hate Speech Explanations on Content ModeratorsTin Nguyen, Jiannan Xu, Aayushi Roy, Hal Daumé III 等EMNLP 2023 · 被引用 1 次
