Improving understandability of feature contributions in model-agnostic explainable AI tools
Sophia Hadash, Martijn C. Willemsen, Chris Snijders, Wijnand A. IJsselsteijn
Abstract
Model-agnostic explainable AI tools explain their predictions by means of ’local’ feature contributions. We empirically investigate two potential improvements over current approaches. The first one is to always present feature contributions in terms of the contribution to the outcome that is perceived as positive by the user (“positive framing”). The second one is to add “semantic labeling”, that explains the directionality of each feature contribution (“this feature leads to +5% eligibility”), reducing additional cognitive processing steps. In a user study, participants evaluated the understandability of explanations for different framing and labeling conditions for loan applications and music recommendations. We found that positive framing improves understandability even when the prediction is negative. Additionally, adding semantic labels eliminates any framing effects on understandability, with positive labels outperforming negative labels. We implemented our suggestions in a package ArgueView[11].
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Understanding the Role of Human Intuition on Reliance in Human-AI Decision-Making with ExplanationsValerie Chen, Q. Vera Liao, Jennifer Wortman Vaughan, Gagan BansalCSCW 2023 · 146 citations
- "As an AI language model, I cannot": Investigating LLM Denials of User RequestsJoel Wester, Tim Schrills, Henning Pohl, Niels van BerkelCHI 2024 · 35 citations
- How Do HCI Researchers Study Cognitive Biases? A Scoping ReviewNattapat Boonprakong, Benjamin Tag, Jorge Gonçalves, Tilman DinglerCHI 2025 · 19 citations
- Interpretability Gone Bad: The Role of Bounded Rationality in How Practitioners Understand Machine LearningHarmanpreet Kaur, Matthew R. Conrad, Davis Rule, Cliff Lampe et al.CSCW 2024 · 14 citations
- Pinning, Sorting, and Categorizing Notifications: A Mixed-methods Usage and Experience Study of Mobile Notification-management FeaturesYong-Han Lin, Li-Ting Su, Uei-Dar Chen, Yi-Chi Lee et al.UbiComp 2024 · 7 citations
Builds on1
Related papers
- The Utility of "Even if" Semifactual Explanation to Optimise Positive OutcomesEoin M. Kenny, Weipeng HuangNeurIPS 2023 · 16 citations
- Causal Shapley Values: Exploiting Causal Knowledge to Explain Individual Predictions of Complex ModelsTom Heskes, Evi Sijben, Ioan Gabriel Bucur, Tom ClaassenNeurIPS 2020 · 235 citations
- ReX: A Framework for Incorporating Temporal Information in Model-Agnostic Local Explanation TechniquesJunhao Liu, Xin ZhangAAAI 2025 · 6 citations
- Impact of Explanation Techniques and Representations on Users' Comprehension and Confidence in Explainable AIJulien Delaunay, Luis Galárraga, Christine Largouët, Niels van BerkelCSCW 2025 · 6 citations
- MaNtLE: Model-agnostic Natural Language ExplainerRakesh R. Menon, Kerem Zaman, Shashank SrivastavaEMNLP 2023 · 1 citation
