Lune

NeurIPS2025Top-tier venue

Approximating Shapley Explanations in Reinforcement Learning

Daniel Beechey, Özgür Simsek

2025Year
1Citations

Abstract

Reinforcement learning has achieved remarkable success in complex decisionmaking environments, yet its lack of transparency limits its deployment in practice, especially in safety-critical settings. Shapley values from cooperative game theory provide a principled framework for explaining reinforcement learning; however, the computational cost of Shapley explanations is an obstacle for their use. We introduce FastSVERL, a scalable method for explaining reinforcement learning by approximating Shapley values. FastSVERL is designed to handle the unique challenges of reinforcement learning, including temporal dependencies across multi-step trajectories, learning from off-policy data, and adapting to evolving agent behaviours in real time. FastSVERL introduces a practical, scalable approach for principled and rigourous interpretability in reinforcement learning.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext c0d2768b-b717-4097-9dd7-9cc892b79081

Builds on5

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines