Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity Constraints
Gabriel Singer, Samuel Gruffaz, Olivier VO VAN, Nicolas Vayatis, Argyris Kalogeratos
摘要
As acquiring reliable ground-truth labels is usually costly, or infeasible, crowdsourcing and aggregation of noisy human annotations is a common alternative. Aggregating subjective labels, though, may amplify individual biases, particularly regarding sensitive features, raising fairness concerns. Nonetheless, fairness in crowdsourced aggregation remains largely unexplored, with no existing convergence guarantees and only limited post-processing approaches for enforcing -fairness under demographic parity. We address this gap by analyzing the fairness of crowdsourced aggregation methods within the -fairness framework, for Majority Vote and Optimal Bayesian aggregation. In the small-crowd regime, we derive an upper bound on the fairness gap of Majority Vote in terms of the fairness gaps of the individual annotators. We further show that the fairness gap of the aggregated consensus converges exponentially fast to that of the ground-truth under interpretable conditions. Since ground-truth itself may still be unfair, we generalize a state-of-the-art multiclass fairness post-processing algorithm from the continuous to the discrete setting, which enforces strict demographic parity constraints on any aggregation rule. Experiments on synthetic and real datasets demonstrate the effectiveness of our approach and corroborate the theoretical insights.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Fair and Optimal Classification via Post-ProcessingRuicheng Xian, Lang Yin, Han ZhaoICML 2023 · 被引用 57 次
- Towards Fair Truth Discovery from Biased Crowdsourced AnswersYanying Li, Haipei Sun, Wendy Hui WangKDD 2020 · 被引用 35 次
- Label Correction of Crowdsourced Noisy Annotations with an Instance-Dependent Noise Transition ModelHui Guo, Boyu Wang, Grace YiNeurIPS 2023 · 被引用 21 次
- Noisy Label Learning with Instance-Dependent Outliers: Identifiability via Crowd WisdomTri Nguyen, Shahana Ibrahim, Xiao FuNeurIPS 2024 · 被引用 14 次
- Learning from Noisy Labels via Conditional Distributionally Robust OptimizationHui Guo, Grace Y. Yi, Boyu WangNeurIPS 2024 · 被引用 8 次
相关 Paper
- Regression under demographic parity constraints via unlabeled post-processingGayane Taturyan, Evgenii Chzhen, Mohamed HebiriNeurIPS 2024 · 被引用 6 次
- Meta Optimality for Demographic Parity Constrained Regression via Post-ProcessingKazuto FukuchiICML 2025
- Fair regression with Wasserstein barycentersEvgenii Chzhen, Christophe Denis, Mohamed Hebiri, Luca Oneto 等NeurIPS 2020 · 被引用 148 次
- Post-hoc bias scoring is optimal for fair classificationWenlong Chen, Yegor Klochkov, Yang LiuICLR 2024 · 被引用 12 次
- Fair regression via plug-in estimator and recalibration with statistical guaranteesEvgenii Chzhen, Christophe Denis, Mohamed Hebiri, Luca Oneto 等NeurIPS 2020 · 被引用 52 次
