Fairness and Explainability: Bridging the Gap towards Fair Model Explanations
Yuying Zhao, Yu Wang, Tyler Derr
摘要
While machine learning models have achieved unprecedented success in real-world applications, they might make biased/unfair decisions for specific demographic groups and hence result in discriminative outcomes. Although research efforts have been devoted to measuring and mitigating bias, they mainly study bias from the result-oriented perspective while neglecting the bias encoded in the decision-making procedure. This results in their inability to capture procedureoriented bias, which therefore limits the ability to have a fully debiasing method. Fortunately, with the rapid development of explainable machine learning, explanations for predictions are now available to gain insights into the procedure. In this work, we bridge the gap between fairness and explainability by presenting a novel perspective of procedure-oriented fairness based on explanations. We identify the procedurebased bias by measuring the gap of explanation quality between different groups with Ratio-based and Value-based Explanation Fairness. The new metrics further motivate us to design an optimization objective to mitigate the procedurebased bias where we observe that it will also mitigate bias from the prediction. Based on our designed optimization objective, we propose a Comprehensive Fairness Algorithm (CFA), which simultaneously fulfills multiple objectives -improving traditional fairness, satisfying explanation fairness, and maintaining the utility performance. Extensive experiments on real-world datasets demonstrate the effectiveness of our proposed CFA and highlight the importance of considering fairness from the explainability perspective. Our code: https://github.com/YuyingZhao/FairExplanations-CFA . Recent years have witnessed the unprecedented success of applying machine learning (ML) models in real-world domains, such as improving the efficiency of information retrieval (Fu et al. 2021; Wang et al. 2022b) and providing convenience with intelligent language translation (Dong et al. 2015) . However, recent studies have revealed that historical data may include patterns of previous discriminatory decisions dominated by sensitive features such as gender, age,
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- PaCEr: Network Embedding From Positional to StructuralYuchen Yan, Yongyi Hu, Qinghai Zhou, Lihui Liu 等WWW 2024 · 被引用 33 次
- FairGP: A Scalable and Fair Graph Transformer Using Graph PartitioningRenqiang Luo, Huafei Huang, Ivan Lee, Chengpei Xu 等AAAI 2025 · 被引用 20 次
- FairGB: A Fair Granular-Ball Generation Method for Data ClassificationQifen Yang, Yuhui Deng, Jiande Huang, Peng Zhou 等ICML 2026
它引用的顶会 Paper4
- Fairness-Aware Explainable Recommendation over Knowledge GraphsZuohui Fu, Yikun Xian, Ruoyuan Gao, Jieyu Zhao 等SIGIR 2020 · 被引用 198 次
- EDITS: Modeling and Mitigating Data Bias for Graph Neural NetworksYushun Dong, Ninghao Liu, Brian Jalaian, Jundong LiWWW 2022 · 被引用 172 次
- Explaining Algorithmic Fairness Through Fairness-Aware Causal Path DecompositionWeishen Pan, Sen Cui, Jiang Bian, Changshui Zhang 等KDD 2021 · 被引用 29 次
- RES: A Robust Framework for Guiding Visual ExplanationYuyang Gao, Tong Steven Sun, Guangji Bai, Siyi Gu 等KDD 2022 · 被引用 29 次
相关 Paper
- Explainable Fairness in RecommendationYingqiang Ge, Juntao Tan, Yan Zhu, Yinglong Xia 等SIGIR 2022 · 被引用 53 次
- AIM: Attributing, Interpreting, Mitigating Data UnfairnessZhining Liu, Ruizhong Qiu, Zhichen Zeng, Yada Zhu 等KDD 2024 · 被引用 4 次
- Constructing Fair Latent Space for Intersection of Fairness and ExplainabilityHyungjun Joo, Hyeonggeun Han, Sehwan Kim, Sangwoo Hong 等AAAI 2025 · 被引用 2 次
- Fairway: a way to build fair ML softwareJoymallya Chakraborty, Suvodeep Majumder, Zhe Yu, Tim MenziesFSE 2020 · 被引用 131 次
- FairFed: Enabling Group Fairness in Federated LearningYahya H. Ezzeldin, Shen Yan, Chaoyang He, Emilio Ferrara 等AAAI 2023 · 被引用 310 次
