Explaining Reinforcement Learning with Shapley Values
Daniel Beechey, Thomas M. S. Smith, Özgür Simsek
摘要
For reinforcement learning systems to be widely adopted, their users must understand and trust them. We present a theoretical analysis of explaining reinforcement learning using Shapley values, following a principled approach from game theory for identifying the contribution of individual players to the outcome of a cooperative game. We call this general framework Shapley Values for Explaining Reinforcement Learning (SVERL). Our analysis exposes the limitations of earlier uses of Shapley values in reinforcement learning. We then develop an approach that uses Shapley values to explain agent performance. In a variety of domains, SVERL produces meaningful explanations that match and supplement human intuition.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Refining Diffusion Planner for Reliable Behavior Synthesis by Automatic Detection of Infeasible PlansKyowoon Lee, Seongun Kim, Jaesik ChoiNeurIPS 2023 · 被引用 31 次
- Towards Multi-dimensional Explanation Alignment for Medical ClassificationLijie Hu, Songning Lai, Wenshuo Chen, Hongru Xiao 等NeurIPS 2024 · 被引用 8 次
- A Comprehensive Study of Shapley Value in Data AnalyticsHong Lin, Shixin Wan, Zhongle Xie, Ke Chen 等VLDB 2025 · 被引用 4 次
- Approximating Shapley Explanations in Reinforcement LearningDaniel Beechey, Özgür SimsekNeurIPS 2025 · 被引用 1 次
- Feature Importance Metrics in the Presence of Missing DataHenrik von Kleist, Joshua Wendland, Ilya Shpitser, Carsten MarrICML 2025
它引用的顶会 Paper2
相关 Paper
- Approximating the Shapley Value without Marginal ContributionsPatrick Kolpaczki, Viktor Bengs, Maximilian Muschalik, Eyke HüllermeierAAAI 2024 · 被引用 43 次
- Causal Shapley Values: Exploiting Causal Knowledge to Explain Individual Predictions of Complex ModelsTom Heskes, Evi Sijben, Ioan Gabriel Bucur, Tom ClaassenNeurIPS 2020 · 被引用 235 次
- Neural Payoff Machines: Predicting Fair and Stable Payoff Allocations Among Team MembersDaphne Cornelisse, Thomas Rood, Yoram Bachrach, Mateusz Malinowski 等NeurIPS 2022 · 被引用 10 次
- SHAQ: Incorporating Shapley Value Theory into Multi-Agent Q-LearningJianhong Wang, Yuan Zhang, Yunjie Gu, Tae-Kyun KimNeurIPS 2022 · 被引用 50 次
- RankSHAP: Shapley Value Based Feature Attributions for Learning to RankTanya Chowdhury, Yair Zick, James AllanICLR 2025
