On Blame Attribution for Accountable Multi-Agent Sequential Decision Making
Stelios Triantafyllou, Adish Singla, Goran Radanovic
Abstract
Blame attribution is one of the key aspects of accountable decision making, as it provides means to quantify the responsibility of an agent for a decision making outcome. In this paper, we study blame attribution in the context of cooperative multi-agent sequential decision making. As a particular setting of interest, we focus on cooperative decision making formalized by Multi-Agent Markov Decision Processes (MMDPs), and we analyze different blame attribution methods derived from or inspired by existing concepts in cooperative game theory. We formalize desirable properties of blame attribution in the setting of interest, and we analyze the relationship between these properties and the studied blame attribution methods. Interestingly, we show that some of the well known blame attribution methods, such as Shapley value, are not performance-incentivizing, while others, such as Banzhaf index, may over-blame agents. To mitigate these value misalignment and fairness issues, we introduce a novel blame attribution method, unique in the set of properties it satisfies, which trade-offs explanatory power (by under-blaming agents) for the aforementioned properties. We further show how to account for uncertainty about agents' decision making policies, and we experimentally: a) validate the qualitative properties of the studied blame attribution methods, and b) analyze their robustness to uncertainty. ... a body of people 1 , holding themselves accountable to nobody, ought not to be trusted by anybody.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4cbf6eee-e297-4420-828e-83c327486735Cited by top-tier papers6
- Game-theoretic Counterfactual Explanation for Graph Neural NetworksChirag Chhablani, Sarthak Jain, Akshay Channesh, Ian A. Kash et al.WWW 2024 · 14 citations
- Performative Reinforcement LearningDebmalya Mandal, Stelios Triantafyllou, Goran RadanovicICML 2023 · 6 citations
- On Corruption-Robustness in Performative Reinforcement LearningVasilis Pollatos, Debmalya Mandal, Goran RadanovicAAAI 2025 · 6 citations
- Agent-Specific Effects: A Causal Effect Propagation Analysis in Multi-Agent MDPsStelios Triantafyllou, Aleksa Sukovic, Debmalya Mandal, Goran RadanovicICML 2024
- Scope Delineation Before Localization: A Two-Stage Framework for Enhancing Failure Attribution in Multi-Agent SystemsKai Sun, Wenqiang Li, Bo Dong, Yuxin Lin et al.AAAI 2026
Builds on3
- Algorithmic Transparency via Quantitative Input Influence: Theory and Experiments with Learning SystemsAnupam Datta, Shayak Sen, Yair ZickS&P 2016 · 774 citations
- Shapley Q-Value: A Local Reward Approach to Solve Global Reward GamesJianhong Wang, Yuan Zhang, Tae-Kyun Kim, Yunjie GuAAAI 2020 · 159 citations
- Human Perceptions on Moral Responsibility of AI: A Case Study in AI-Assisted Bail Decision-MakingGabriel Lima, Nina Grgic-Hlaca, Meeyoung ChaCHI 2021 · 71 citations
Related papers
- Find True Collaborators: Banzhaf Index-based Cross View Alignment for Partially View-aligned ClusteringShanghui Deng, Xiao Zheng, Chang Tang, Kun Sun et al.ACM MM 2025 · 2 citations
- Counterfactual Effect Decomposition in Multi-Agent Sequential Decision MakingStelios Triantafyllou, Aleksa Sukovic, Yasaman Zolfimoselo, Goran RadanovicICML 2025
- Responsibility Attribution in Parameterized Markovian ModelsChristel Baier, Florian Funke, Rupak MajumdarAAAI 2021 · 9 citations
- Backward Responsibility in Transition Systems Using General Power IndicesChristel Baier, Roxane van den Bossche, Sascha Klüppelholz, Johannes Lehmann et al.AAAI 2024 · 4 citations
- Energy-Based Learning for Cooperative Games, with Applications to Valuation Problems in Machine LearningYatao Bian, Yu Rong, Tingyang Xu, Jiaxiang Wu et al.ICLR 2022 · 17 citations
