Aleatoric and Epistemic Discrimination: Fundamental Limits of Fairness Interventions
Hao Wang, Luxi He, Rui Gao, Flávio P. Calmon
摘要
Machine learning (ML) models can underperform on certain population groups due to choices made during model development and bias inherent in the data. We categorize sources of discrimination in the ML pipeline into two classes: aleatoric discrimination, which is inherent in the data distribution, and epistemic discrimination, which is due to decisions made during model development. We quantify aleatoric discrimination by determining the performance limits of a model under fairness constraints, assuming perfect knowledge of the data distribution. We demonstrate how to characterize aleatoric discrimination by applying Blackwell's results on comparing statistical experiments. We then quantify epistemic discrimination as the gap between a model's accuracy when fairness constraints are applied and the limit posed by aleatoric discrimination. We apply this approach to benchmark existing fairness interventions and investigate fairness risks in data with missing values. Our results indicate that state-of-the-art fairness interventions are effective at removing epistemic discrimination on standard (overused) tabular datasets. However, when data has missing values, there is still significant room for improvement in handling aleatoric discrimination.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Demystifying Local & Global Fairness Trade-offs in Federated Learning Using Partial Information DecompositionFaisal Hamman, Sanghamitra DuttaICLR 2024 · 被引用 9 次
- Debiasing Attention Mechanism in Transformer without DemographicsShenyu Lu, Yipei Wang, Xiaoqian WangICLR 2024 · 被引用 7 次
- Achievable Fairness on Your Data With Utility GuaranteesMuhammad Faaiz Taufiq, Jean-Francois Ton, Yang LiuNeurIPS 2024 · 被引用 3 次
- Efficient Fairness-Performance Pareto Front ComputationMark Kozdoba, Binyamin Perets, Shie MannorNeurIPS 2025 · 被引用 2 次
- Disparate Conditional Prediction in Multiclass ClassifiersSivan Sabato, Eran Treister, Elad Yom-TovICML 2025
它引用的顶会 Paper13
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 被引用 671 次
- Minimax Pareto Fairness: A Multi Objective PerspectiveNatalia Martínez, Martín Bertrán, Guillermo SapiroICML 2020 · 被引用 232 次
- Is There a Trade-Off Between Fairness and Accuracy? A Perspective Using Mismatched Hypothesis TestingSanghamitra Dutta, Dennis Wei, Hazar Yueksel, Pin-Yu Chen 等ICML 2020 · 被引用 171 次
- Robust Optimization for Fairness with Noisy Protected GroupsSerena Lutong Wang, Wenshuo Guo, Harikrishna Narasimhan, Andrew Cotter 等NeurIPS 2020 · 被引用 134 次
- Fairness with Overlapping Groups; a Probabilistic PerspectiveForest Yang, Mouhamadou Cisse, Oluwasanmi KoyejoNeurIPS 2020 · 被引用 71 次
相关 Paper
- Fairness without Imputation: A Decision Tree Approach for Fair Prediction with Missing ValuesHaewon Jeong, Hao Wang, Flávio P. CalmonAAAI 2022 · 被引用 48 次
- Rethinking Aleatoric and Epistemic UncertaintyFreddie Bickford Smith, Jannik Kossen, Eleanor Trollope, Mark van der Wilk 等ICML 2025
- From Risk to Uncertainty: Generating Predictive Uncertainty Measures via Bayesian EstimationNikita Kotelevskii, Vladimir Kondratyev, Martin Takác, Eric Moulines 等ICLR 2025
- Assessing Fairness in the Presence of Missing DataYiliang Zhang, Qi LongNeurIPS 2021 · 被引用 51 次
- Adapting Fairness Interventions to Missing ValuesRaymond Feng, Flávio P. Calmon, Hao WangNeurIPS 2023 · 被引用 20 次
