Adapting Fairness Interventions to Missing Values
Raymond Feng, Flávio P. Calmon, Hao Wang
摘要
Missing values in real-world data pose a significant and unique challenge to algorithmic fairness. Different demographic groups may be unequally affected by missing data, and the standard procedure for handling missing values where first data is imputed, then the imputed data is used for classification -- a procedure referred to as"impute-then-classify"-- can exacerbate discrimination. In this paper, we analyze how missing values affect algorithmic fairness. We first prove that training a classifier from imputed data can significantly worsen the achievable values of group fairness and average accuracy. This is because imputing data results in the loss of the missing pattern of the data, which often conveys information about the predictive label. We present scalable and adaptive algorithms for fair classification with missing values. These algorithms can be combined with any preexisting fairness-intervention algorithm to handle all possible missing patterns while preserving information encoded within the missing patterns. Numerical experiments with state-of-the-art fairness interventions demonstrate that our adaptive algorithms consistently achieve higher fairness and accuracy than impute-then-classify across different datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- On the Maximal Local Disparity of Fairness-Aware ClassifiersJinqiu Jin, Haoxuan Li, Fuli FengICML 2024 · 被引用 5 次
- Still More Shades of Null: An Evaluation Suite for Responsible Missing Value Imputation [Experiment, Analysis and Benchmark]Falaah Arif Khan, Denys Herasymuk, Nazar Protsiv, Julia StoyanovichVLDB 2025 · 被引用 3 次
- Addressing Missing Data Issue for Diffusion-based RecommendationWenyu Mao, Zhengyi Yang, Jiancan Wu, Haozhe Liu 等SIGIR 2025 · 被引用 2 次
- Fair Graph Machine Learning under Adversarial Missingness ProcessesDebolina Halder Lina, Arlei SilvaICLR 2026
- An Effective Theory of Bias AmplificationArjun Subramonian, Samuel J. Bell, Levent Sagun, Elvis DohmatobICLR 2025
它引用的顶会 Paper10
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 被引用 671 次
- Predictive Multiplicity in ClassificationCharles T. Marx, Flávio P. Calmon, Berk UstunICML 2020 · 被引用 197 次
- What's a good imputation to predict with missing values?Marine Le Morvan, Julie Josse, Erwan Scornet, Gaël VaroquauxNeurIPS 2021 · 被引用 95 次
- Fairness with Overlapping Groups; a Probabilistic PerspectiveForest Yang, Mouhamadou Cisse, Oluwasanmi KoyejoNeurIPS 2020 · 被引用 71 次
- Can I Trust My Fairness Metric? Assessing Fairness with Unlabeled Data and Bayesian InferenceDisi Ji, Padhraic Smyth, Mark SteyversNeurIPS 2020 · 被引用 57 次
相关 Paper
- Fairness without Imputation: A Decision Tree Approach for Fair Prediction with Missing ValuesHaewon Jeong, Hao Wang, Flávio P. CalmonAAAI 2022 · 被引用 48 次
- Fairness-Aware Classification over Incomplete DataXiaoye Miao, Lei Qiang, Guilin Huang, Yangyang Wu 等SIGIR 2025 · 被引用 1 次
- Assessing Fairness in the Presence of Missing DataYiliang Zhang, Qi LongNeurIPS 2021 · 被引用 51 次
- Multigroup RobustnessLunjia Hu, Charlotte Peale, Judy Hanwen ShenICML 2024 · 被引用 2 次
- The Importance of Modeling Data Missingness in Algorithmic Fairness: A Causal PerspectiveNaman Goel, Alfonso Amayuelas, Amit Deshpande, Amit SharmaAAAI 2021 · 被引用 36 次
