Statistical inference for individual fairness
Subha Maity, Songkai Xue, Mikhail Yurochkin, Yuekai Sun
摘要
As we rely on machine learning (ML) models to make more consequential decisions, the issue of ML models perpetuating or even exacerbating undesirable historical biases (e.g. gender and racial biases) has come to the fore of the public's attention. In this paper, we focus on the problem of detecting violations of individual fairness in ML models. We formalize the problem as measuring the susceptibility of ML models against a form of adversarial attack and develop a suite of inference tools for the adversarial cost function. The tools allow auditors to assess the individual fairness of ML models in a statistically-principled way: form confidence intervals for the worst-case performance differential between similar individuals and test hypotheses of model fairness with (asymptotic) non-coverage/Type I error rate control. We demonstrate the utility of our tools in a real-world case study 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Fairness Evaluation in Text Classification: Machine Learning Practitioner Perspectives of Individual and Group FairnessZahra Ashktorab, Benjamin Hoover, Mayank Agarwal, Casey Dugan 等CHI 2023 · 被引用 13 次
- Learning Antidote Data to Individual UnfairnessPeizhao Li, Ethan Xia, Hongfu LiuICML 2023 · 被引用 11 次
- Empirical Likelihood for Fair ClassificationPangpang Liu, Yichuan ZhaoICLR 2024 · 被引用 1 次
它引用的顶会 Paper2
相关 Paper
- Active Fourier Auditor for Estimating Distributional Properties of ML ModelsAyoub Ajarra, Bishwamittra Ghosh, Debabrota BasuAAAI 2025 · 被引用 5 次
- Online Fairness Auditing through Iterative RefinementPranav Maneriker, Codi Burley, Srinivasan ParthasarathyKDD 2023 · 被引用 6 次
- Approximation-guided Fairness Testing through Discriminatory Space AnalysisZhenjiang Zhao, Takahisa Toda, Takashi KitamuraASE 2024
- SenSeI: Sensitive Set Invariance for Enforcing Individual FairnessMikhail Yurochkin, Yuekai SunICLR 2021 · 被引用 53 次
- MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient SearchZhaohui Wang, Min Zhang, Jingran Yang, Bojie Shao 等ICSE 2024 · 被引用 6 次
