A Game-Theoretic Framework for Measuring and Explaining Metric Compatibility in Fair Machine Learning
Lingfeng Zhang, Jingran Yang, Zhaohui Wang, Min Zhang, Qing Zhang
摘要
Machine learning fairness research documents trade-offs but lacks quantitative frameworks to measure intrinsic metric compatibility without requiring causal graphs. We introduce a game-theoretic framework that decomposes metrics into interaction vectors, enabling compatibility measurement between metrics via cosine similarity and mechanistic attribution to attribute coalitions. Through analysis of 6 datasets, 7 models, and 6 debiasing methods, we reveal that fairness and utility are often structurally orthogonal (median compatibility ) rather than diametrically opposed, with conflicts driven by sparse, low-order interactions. We further show that debiasing improves fairness by compressing the compatibility space—reducing compatibility of both synergistic and conflicting relationships—rather than eliminating conflicts, providing a mechanistic basis for understanding metric alignment.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 被引用 2,148 次
- Bias in machine learning software: why? how? what to do?Joymallya Chakraborty, Suvodeep Majumder, Tim MenziesFSE 2021 · 被引用 186 次
- Interpreting Multivariate Shapley Interactions in DNNsHao Zhang, Yichen Xie, Longjie Zheng, Die Zhang 等AAAI 2021 · 被引用 70 次
- MAAT: a novel ensemble approach to addressing fairness and performance bugs for machine learning softwareZhenpeng Chen, Jie M. Zhang, Federica Sarro, Mark HarmanFSE 2022 · 被引用 65 次
- Does a Neural Network Really Encode Symbolic Concepts?Mingjie Li, Quanshi ZhangICML 2023 · 被引用 35 次
相关 Paper
- Information-Theoretic Bias Reduction via Causal View of Spurious CorrelationSeonguk Seo, Joon-Young Lee, Bohyung HanAAAI 2022 · 被引用 29 次
- Understanding Fairness and Prediction Error through Subspace Decomposition and Influence AnalysisEnze Shi, Pankaj Bhagwat, Zhixian Yang, Linglong Kong 等NeurIPS 2025
- Fairness via Independence: A General Regularization Framework for Machine LearningYezi Liu, Hanning Chen, Wenjun Huang, Yang Ni 等ICLR 2026
- Causal Disentangled Anchor Learning for Scalable Fair Multi-view ClusteringSuyuan Liu, Shengfei Wei, Wenjing Yang, Shengju Yu 等ICML 2026
- Towards Attributions of Input Variables in a CoalitionXinhao Zheng, Huiqi Deng, Quanshi ZhangICML 2025
