Bounding the fairness and accuracy of classifiers from population statistics
Sivan Sabato, Elad Yom-Tov
摘要
We consider the study of a classification model whose properties are impossible to estimate using a validation set, either due to the absence of such a set or because access to the classifier, even as a black-box, is impossible. Instead, only aggregate statistics on the rate of positive predictions in each of several sub-populations are available, as well as the true rates of positive labels in each of these sub-populations. We show that these aggregate statistics can be used to lowerbound the discrepancy of a classifier, which is a measure that balances inaccuracy and unfairness. To this end, we define a new measure of unfairness, equal to the fraction of the population on which the classifier behaves differently, compared to its global, ideally fair behavior, as defined by the measure of equalized odds. We propose an efficient and practical procedure for finding the best possible lower bound on the discrepancy of the classifier, given the aggregate statistics, and demonstrate in experiments the empirical tightness of this lower bound, as well as its possible uses on various types of problems, ranging from estimating the quality of voting polls to measuring the effectiveness of patient identification from internet search queries. The code and data are available at https://github.com/ sivansabato/bfa .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Is There a Trade-Off Between Fairness and Accuracy? A Perspective Using Mismatched Hypothesis TestingSanghamitra Dutta, Dennis Wei, Hazar Yueksel, Pin-Yu Chen 等ICML 2020 · 被引用 171 次
- Are My Deep Learning Systems Fair? An Empirical Study of Fixed-Seed TrainingShangshu Qian, Hung Viet Pham, Thibaud Lutellier, Zeou Hu 等NeurIPS 2021 · 被引用 49 次
- Demystifying Local & Global Fairness Trade-offs in Federated Learning Using Partial Information DecompositionFaisal Hamman, Sanghamitra DuttaICLR 2024 · 被引用 9 次
- Can Information Flows Suggest Targets for Interventions in Neural Circuits?Praveen Venkatesh, Sanghamitra Dutta, Neil Ashim Mehta, Pulkit GroverNeurIPS 2021 · 被引用 8 次
- Disparate Conditional Prediction in Multiclass ClassifiersSivan Sabato, Eran Treister, Elad Yom-TovICML 2025
相关 Paper
- Achieving Equalized Odds by Resampling Sensitive AttributesYaniv Romano, Stephen Bates, Emmanuel J. CandèsNeurIPS 2020 · 被引用 65 次
- The Price of Fairness in Active Learning: Fundamental Limits and Optimal Label AcquisitionChang Lu, Yizheng ZhaoKDD 2026
- Estimating Structural Disparities for Face ModelsShervin Ardeshir, Cristina Segalin, Nathan KallusCVPR 2022 · 被引用 2 次
- Estimating and Controlling for Equalized Odds via Sensitive Attribute PredictorsBeepul Bharti, Paul H. Yi, Jeremias SulamNeurIPS 2023 · 被引用 8 次
- Mitigating Source Bias for Fairer Weak SupervisionChangho Shin, Sonia Cromp, Dyah Adila, Frederic SalaNeurIPS 2023 · 被引用 5 次
