FAIRER: Fairness as Decision Rationale Alignment
Tianlin Li, Qing Guo, Aishan Liu, Mengnan Du, Zhiming Li, Yang Liu
摘要
Deep neural networks (DNNs) have made significant progress, but often suffer from fairness issues, as deep models typically show distinct accuracy differences among certain subgroups (e.g., males and females). Existing research addresses this critical issue by employing fairness-aware loss functions to constrain the last-layer outputs and directly regularize DNNs. Although the fairness of DNNs is improved, it is unclear how the trained network makes a fair prediction, which limits future fairness improvements. In this paper, we investigate fairness from the perspective of decision rationale and define the parameter parity score to characterize the fair decision process of networks by analyzing neuron influence in various subgroups. Extensive empirical studies show that the unfair issue could arise from the unaligned decision rationales of subgroups. Existing fairness regularization terms fail to achieve decision rationale alignment because they only constrain last-layer outputs while ignoring intermediate neuron alignment. To address the issue, we formulate the fairness as a new task, i.e., decision rationale alignment that requires DNNs' neurons to have consistent responses on subgroups at both intermediate processes and the final prediction. To make this idea practical during optimization, we relax the naive objective function and propose gradient-guided parity alignment, which encourages gradient-weighted consistency of neurons across subgroups. Extensive experiments on a variety of datasets show that our method can significantly enhance fairness while sustaining a high level of accuracy and outperforming other approaches by a wide margin.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- FedMut: Generalized Federated Learning via Stochastic MutationMing Hu, Yue Cao, Anran Li, Zhiming Li 等AAAI 2024 · 被引用 46 次
- Is Aggregation the Only Choice? Federated Learning via Layer-wise Model RecombinationMing Hu, Zhihao Yue, Xiaofei Xie, Cheng Chen 等KDD 2024 · 被引用 20 次
- RUNNER: Responsible UNfair NEuron Repair for Enhancing Deep Neural Network FairnessTianlin Li, Yue Cao, Jian Zhang, Shiqian Zhao 等ICSE 2024 · 被引用 11 次
- MultiSFL: Towards Accurate Split Federated Learning via Multi-Model Aggregation and Knowledge ReplayZeke Xia, Ming Hu, Dengke Yan, Ruixuan Liu 等AAAI 2025 · 被引用 8 次
- GenderCARE: A Comprehensive Framework for Assessing and Reducing Gender Bias in Large Language ModelsKunsheng Tang, Wenbo Zhou, Jie Zhang, Aishan Liu 等CCS 2024 · 被引用 7 次
它引用的顶会 Paper12
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang 等ICCV 2019 · 被引用 469 次
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong 等ICSE 2020 · 被引用 127 次
- Fairness via Representation NeutralizationMengnan Du, Subhabrata Mukherjee, Guanchu Wang, Ruixiang Tang 等NeurIPS 2021 · 被引用 91 次
- Controllable Guarantees for Fair Outcomes via Contrastive Information EstimationUmang Gupta, Aaron M. Ferber, Bistra Dilkina, Greg Ver SteegAAAI 2021 · 被引用 78 次
- NeuronFair: Interpretable White-Box Fairness Testing through Biased Neuron IdentificationHaibin Zheng, Zhiqing Chen, Tianyu Du, Xuhong Zhang 等ICSE 2022 · 被引用 58 次
相关 Paper
- Fairneuron: Improving Deep Neural Network Fairness with Adversary Games on Selective NeuronsXuanqi Gao, Juan Zhai, Shiqing Ma, Chao Shen 等ICSE 2022 · 被引用 36 次
- Adaptive fairness improvement based on causality analysisMengdi Zhang, Jun SunFSE 2022 · 被引用 35 次
- Are Two Heads the Same as One? Identifying Disparate Treatment in Fair Neural NetworksMichael Lohaus, Matthäus Kleindessner, Krishnaram Kenthapadi, Francesco Locatello 等NeurIPS 2022 · 被引用 15 次
- Towards Debiasing DNN Models from Spurious Feature InfluenceMengnan Du, Ruixiang Tang, Weijie Fu, Xia HuAAAI 2022 · 被引用 9 次
- Some Optimizers are More Equal: Understanding the Role of Optimizers in Group FairnessMojtaba Kolahdouzi, Hatice Gunes, Ali EtemadNeurIPS 2025
