FAIRER: Fairness as Decision Rationale Alignment
Tianlin Li, Qing Guo, Aishan Liu, Mengnan Du, Zhiming Li, Yang Liu
Abstract
Deep neural networks (DNNs) have made significant progress, but often suffer from fairness issues, as deep models typically show distinct accuracy differences among certain subgroups (e.g., males and females). Existing research addresses this critical issue by employing fairness-aware loss functions to constrain the last-layer outputs and directly regularize DNNs. Although the fairness of DNNs is improved, it is unclear how the trained network makes a fair prediction, which limits future fairness improvements. In this paper, we investigate fairness from the perspective of decision rationale and define the parameter parity score to characterize the fair decision process of networks by analyzing neuron influence in various subgroups. Extensive empirical studies show that the unfair issue could arise from the unaligned decision rationales of subgroups. Existing fairness regularization terms fail to achieve decision rationale alignment because they only constrain last-layer outputs while ignoring intermediate neuron alignment. To address the issue, we formulate the fairness as a new task, i.e., decision rationale alignment that requires DNNs' neurons to have consistent responses on subgroups at both intermediate processes and the final prediction. To make this idea practical during optimization, we relax the naive objective function and propose gradient-guided parity alignment, which encourages gradient-weighted consistency of neurons across subgroups. Extensive experiments on a variety of datasets show that our method can significantly enhance fairness while sustaining a high level of accuracy and outperforming other approaches by a wide margin.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 01c70499-78f7-45c8-9916-65bd6cc566b0Cited by top-tier papers17
- FedMut: Generalized Federated Learning via Stochastic MutationMing Hu, Yue Cao, Anran Li, Zhiming Li et al.AAAI 2024 · 46 citations
- Is Aggregation the Only Choice? Federated Learning via Layer-wise Model RecombinationMing Hu, Zhihao Yue, Xiaofei Xie, Cheng Chen et al.KDD 2024 · 20 citations
- RUNNER: Responsible UNfair NEuron Repair for Enhancing Deep Neural Network FairnessTianlin Li, Yue Cao, Jian Zhang, Shiqian Zhao et al.ICSE 2024 · 11 citations
- MultiSFL: Towards Accurate Split Federated Learning via Multi-Model Aggregation and Knowledge ReplayZeke Xia, Ming Hu, Dengke Yan, Ruixuan Liu et al.AAAI 2025 · 8 citations
- GenderCARE: A Comprehensive Framework for Assessing and Reducing Gender Bias in Large Language ModelsKunsheng Tang, Wenbo Zhou, Jie Zhang, Aishan Liu et al.CCS 2024 · 7 citations
Builds on12
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang et al.ICCV 2019 · 469 citations
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong et al.ICSE 2020 · 127 citations
- Fairness via Representation NeutralizationMengnan Du, Subhabrata Mukherjee, Guanchu Wang, Ruixiang Tang et al.NeurIPS 2021 · 91 citations
- Controllable Guarantees for Fair Outcomes via Contrastive Information EstimationUmang Gupta, Aaron M. Ferber, Bistra Dilkina, Greg Ver SteegAAAI 2021 · 78 citations
- NeuronFair: Interpretable White-Box Fairness Testing through Biased Neuron IdentificationHaibin Zheng, Zhiqing Chen, Tianyu Du, Xuhong Zhang et al.ICSE 2022 · 58 citations
Related papers
- Fairneuron: Improving Deep Neural Network Fairness with Adversary Games on Selective NeuronsXuanqi Gao, Juan Zhai, Shiqing Ma, Chao Shen et al.ICSE 2022 · 36 citations
- Adaptive fairness improvement based on causality analysisMengdi Zhang, Jun SunFSE 2022 · 35 citations
- Are Two Heads the Same as One? Identifying Disparate Treatment in Fair Neural NetworksMichael Lohaus, Matthäus Kleindessner, Krishnaram Kenthapadi, Francesco Locatello et al.NeurIPS 2022 · 15 citations
- Towards Debiasing DNN Models from Spurious Feature InfluenceMengnan Du, Ruixiang Tang, Weijie Fu, Xia HuAAAI 2022 · 9 citations
- Some Optimizers are More Equal: Understanding the Role of Optimizers in Group FairnessMojtaba Kolahdouzi, Hatice Gunes, Ali EtemadNeurIPS 2025
