Experimental Analysis of Multi-Step Pipelines for Fair Classifications - More than the Sum of Their Parts?
Nico Lässig, Melanie Herschel
摘要
The problem of biased machine learning predictions has led to many alternative approaches to mitigate the problem. They are typically studied and evaluated by focusing on the input data, the trained model, and the performance of the model predictions. We take a broader perspective, considering approaches in the context of a multi-step pipeline. We study fair classification in a pipeline comprising multiple data preparation steps, parameter optimization, and three types of approaches (pre-, in-, and post-processing) designed to reduce bias that may be applied consecutively. This pipeline leads to a trained model to be evaluated in terms of quality (e.g., accuracy) and fairness. We experimentally evaluate the effect differently combined implementations of the pipeline components have on the performance of more than 40 fairness-inducing algorithms. Key findings made possible by this pipeline perspective include: (1) Choosing a bias reducing algorithm greatly simplifies when implementing suited data preparation or parameter optimization, as the difference in performance between methods shrinks, making almost any choice a good one. (2) Several component or pipeline implementations often assumed to have positive or negative effects on performance prove to have little or even contrary effects to the expectations. (3) While many approaches have been published for fair classification in the last decade and shown to improve on previous solutions in specific settings, our broad analysis reveals a stagnating performance trend. Our analysis shows that synergetic effects between pipeline components need to be carefully taken into account for further research on fair end-to-end data processing. It further raises the more fundamental question of how the study of the problem evolves, both in terms of proposed solutions and benchmarking.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- FRAPPÉ: A Group Fairness Framework for Post-Processing EverythingAlexandru Tifrea, Preethi Lahoti, Ben Packer, Yoni Halpern 等ICML 2024 · 被引用 15 次
- Fair preprocessing: towards understanding compositional fairness of data transformers in machine learning pipelineSumon Biswas, Hridesh RajanFSE 2021 · 被引用 101 次
- Demystifying the Optimal Fair Classifier in Multi-Class ClassificationLi Zhang, Yuyuan Li, XiaoHua Feng, Jiaming Zhang 等ICML 2026
- Through the Data Management Lens: Experimental Analysis and Evaluation of Fair ClassificationMaliha Tashfia Islam, Anna Fariha, Alexandra Meliou, Babak SalimiSIGMOD 2022 · 被引用 29 次
- Beyond Adult and COMPAS: Fair Multi-Class Prediction via Information ProjectionWael Alghamdi, Hsiang Hsu, Haewon Jeong, Hao Wang 等NeurIPS 2022 · 被引用 57 次
