A Game-Theoretic Framework for Measuring and Explaining Metric Compatibility in Fair Machine Learning
Lingfeng Zhang, Jingran Yang, Zhaohui Wang, Min Zhang, Qing Zhang
Abstract
Machine learning fairness research documents trade-offs but lacks quantitative frameworks to measure intrinsic metric compatibility without requiring causal graphs. We introduce a game-theoretic framework that decomposes metrics into interaction vectors, enabling compatibility measurement between metrics via cosine similarity and mechanistic attribution to attribute coalitions. Through analysis of 6 datasets, 7 models, and 6 debiasing methods, we reveal that fairness and utility are often structurally orthogonal (median compatibility ) rather than diametrically opposed, with conflicts driven by sparse, low-order interactions. We further show that debiasing improves fairness by compressing the compatibility space—reducing compatibility of both synergistic and conflicting relationships—rather than eliminating conflicts, providing a mechanistic basis for understanding metric alignment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on19
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 2,148 citations
- Bias in machine learning software: why? how? what to do?Joymallya Chakraborty, Suvodeep Majumder, Tim MenziesFSE 2021 · 186 citations
- Interpreting Multivariate Shapley Interactions in DNNsHao Zhang, Yichen Xie, Longjie Zheng, Die Zhang et al.AAAI 2021 · 70 citations
- MAAT: a novel ensemble approach to addressing fairness and performance bugs for machine learning softwareZhenpeng Chen, Jie M. Zhang, Federica Sarro, Mark HarmanFSE 2022 · 65 citations
- Does a Neural Network Really Encode Symbolic Concepts?Mingjie Li, Quanshi ZhangICML 2023 · 35 citations
Related papers
- Information-Theoretic Bias Reduction via Causal View of Spurious CorrelationSeonguk Seo, Joon-Young Lee, Bohyung HanAAAI 2022 · 29 citations
- Understanding Fairness and Prediction Error through Subspace Decomposition and Influence AnalysisEnze Shi, Pankaj Bhagwat, Zhixian Yang, Linglong Kong et al.NeurIPS 2025
- Fairness via Independence: A General Regularization Framework for Machine LearningYezi Liu, Hanning Chen, Wenjun Huang, Yang Ni et al.ICLR 2026
- Causal Disentangled Anchor Learning for Scalable Fair Multi-view ClusteringSuyuan Liu, Shengfei Wei, Wenjing Yang, Shengju Yu et al.ICML 2026
- Towards Attributions of Input Variables in a CoalitionXinhao Zheng, Huiqi Deng, Quanshi ZhangICML 2025
