Training individually fair ML models with sensitive subspace robustness
Mikhail Yurochkin, Amanda Bower, Yuekai Sun
Abstract
We propose an approach to training machine learning models that are fair in the sense that their performance is invariant under certain perturbations to the features. For example, the performance of a resume screening system should be invariant under changes to the name of the applicant. We formalize this intuitive notion of fairness by connecting it to the original notion of individual fairness put forth by Dwork et al and show that the proposed approach achieves this notion of fairness. We also demonstrate the effectiveness of the approach on two machine learning tasks that are susceptible to gender and racial biases.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers41
- Exactly Computing the Local Lipschitz Constant of ReLU NetworksMatt Jordan, Alexandros G. DimakisNeurIPS 2020 · 156 citations
- Post-processing for Individual FairnessFelix Petersen, Debarghya Mukherjee, Yuekai Sun, Mikhail YurochkinNeurIPS 2021 · 115 citations
- Learning Certified Individually Fair RepresentationsAnian Ruoss, Mislav Balunovic, Marc Fischer, Martin T. VechevNeurIPS 2020 · 112 citations
- Fast Model DeBias with Machine UnlearningRuizhe Chen, Jianfei Yang, Huimin Xiong, Jianhong Bai et al.NeurIPS 2023 · 110 citations
- Two Simple Ways to Learn Individual Fairness Metrics from DataDebarghya Mukherjee, Mikhail Yurochkin, Moulinath Banerjee, Yuekai SunICML 2020 · 109 citations
Builds on1
Related papers
- SenSeI: Sensitive Set Invariance for Enforcing Individual FairnessMikhail Yurochkin, Yuekai SunICLR 2021 · 53 citations
- On the Alignment between Fairness and Accuracy: from the Perspective of Adversarial RobustnessJunyi Chai, Taeuk Jang, Jing Gao, Xiaoqian WangICML 2025
- Robust Fairness Under Covariate ShiftAshkan Rezaei, Anqi Liu, Omid Memarrast, Brian D. ZiebartAAAI 2021 · 94 citations
- Correct-by-Construction: Certified Individual Fairness through Neural Network TrainingRuihan Zhang, Jun SunOOPSLA 2025 · 1 citation
- Learning Antidote Data to Individual UnfairnessPeizhao Li, Ethan Xia, Hongfu LiuICML 2023 · 11 citations
