How Far Can Fairness Constraints Help Recover From Biased Data?
Mohit Sharma, Amit Deshpande
摘要
A general belief in fair classification is that fairness constraints incur a trade-off with accuracy, which biased data may worsen. Contrary to this belief, Blum&Stangl (2019) show that fair classification with equal opportunity constraints even on extremely biased data can recover optimally accurate and fair classifiers on the original data distribution. Their result is interesting because it demonstrates that fairness constraints can implicitly rectify data bias and simultaneously overcome a perceived fairness-accuracy trade-off. Their data bias model simulates under-representation and label bias in underprivileged population, and they show the above result on a stylized data distribution with i.i.d. label noise, under simple conditions on the data distribution and bias parameters. We propose a general approach to extend the result of Blum&Stangl (2019) to different fairness constraints, data bias models, data distributions, and hypothesis classes. We strengthen their result, and extend it to the case when their stylized distribution has labels with Massart noise instead of i.i.d. noise. We prove a similar recovery result for arbitrary data distributions using fair reject option classifiers. We further generalize it to arbitrary data distributions and arbitrary hypothesis classes, i.e., we prove that for any data distribution, if the optimally accurate classifier in a given hypothesis class is fair and robust, then it can be recovered through fair classification with equal opportunity constraints on the biased distribution whenever the bias parameters satisfy certain simple conditions. Finally, we show applications of our technique to time-varying data bias in classification and fair machine learning pipelines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- On Optimal Steering to Achieve Exact FairnessMohit Sharma, Amit Deshpande, Chiranjib Bhattacharyya, Rajiv Ratn ShahNeurIPS 2025 · 被引用 2 次
- Bias In, Bias Out? Finding Unbiased Subnetworks in Vanilla ModelsIvan Luiz De Moura Matos, Abdel Djalil Sad Saoud, Ekaterina Lakovleva, Vito Paolo Pastore 等CVPR 2026
- A Game-Theoretic Framework for Measuring and Explaining Metric Compatibility in Fair Machine LearningLingfeng Zhang, Jingran Yang, Zhaohui Wang, Min Zhang 等ICML 2026
它引用的顶会 Paper8
- Is There a Trade-Off Between Fairness and Accuracy? A Perspective Using Mismatched Hypothesis TestingSanghamitra Dutta, Dennis Wei, Hazar Yueksel, Pin-Yu Chen 等ICML 2020 · 被引用 171 次
- Classification with Rejection Based on Cost-sensitive ClassificationNontawat Charoenphakdee, Zhenghang Cui, Yivan Zhang, Masashi SugiyamaICML 2021 · 被引用 78 次
- Fair Classification with Noisy Protected Attributes: A Framework with Provable GuaranteesL. Elisa Celis, Lingxiao Huang, Vijay Keswani, Nisheeth K. VishnoiICML 2021 · 被引用 67 次
- Does enforcing fairness mitigate biases caused by subpopulation shift?Subha Maity, Debarghya Mukherjee, Mikhail Yurochkin, Yuekai SunNeurIPS 2021 · 被引用 32 次
- Through the Data Management Lens: Experimental Analysis and Evaluation of Fair ClassificationMaliha Tashfia Islam, Anna Fariha, Alexandra Meliou, Babak SalimiSIGMOD 2022 · 被引用 29 次
相关 Paper
- On the Impossibility of Non-trivial Accuracy in Presence of Fairness ConstraintsCarlos Pinzón, Catuscia Palamidessi, Pablo Piantanida, Frank ValenciaAAAI 2022 · 被引用 11 次
- Optimal Fair Learning Robust to Adversarial Distribution ShiftSushant Agarwal, Amit Deshpande, Rajmohan Rajaraman, Ravi SundaramICML 2025
- Demystifying the Optimal Fair Classifier in Multi-Class ClassificationLi Zhang, Yuyuan Li, XiaoHua Feng, Jiaming Zhang 等ICML 2026
- FaiREE: fair classification with finite-sample and distribution-free guaranteePuheng Li, James Zou, Linjun ZhangICLR 2023
- FIFA: Making Fairness More Generalizable in Classifiers Trained on Imbalanced DataZhun Deng, Jiayao Zhang, Linjun Zhang, Ting Ye 等ICLR 2023 · 被引用 7 次
