A Large-Scale Empirical Study on Improving the Fairness of Image Classification Models
Junjie Yang, Jiajun Jiang, Zeyu Sun, Junjie Chen
摘要
Fairness has been a critical issue that affects the adoption of deep learning models in real practice. To improve model fairness, many existing methods have been proposed and evaluated to be effective in their own contexts. However, there is still no systematic evaluation among them for a comprehensive comparison under the same context, which makes it hard to understand the performance distinction among them, hindering the research progress and practical adoption of them. To fill this gap, this paper endeavours to conduct the first large-scale empirical study to comprehensively compare the performance of existing state-of-the-art fairness improving techniques. Specifically, we target the widely-used application scenario of image classification, and utilized three different datasets and five commonly-used performance metrics to assess in total 13 methods from diverse categories. Our findings reveal substantial variations in the performance of each method across different datasets and sensitive attributes, indicating over-fitting on specific datasets by many existing methods. Furthermore, different fairness evaluation metrics, due to their distinct focuses, yield significantly different assessment results. Overall, we observe that pre-processing methods and in-processing methods outperform post-processing methods, with pre-processing methods exhibiting the best performance. Our empirical study offers comprehensive recommendations for enhancing fairness in deep learning models. We approach the problem from multiple dimensions, aiming to provide a uniform evaluation platform and inspire researchers to explore more effective fairness solutions via a set of implications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?Yunbo Lyu, Zhou Yang, Yuqing Niu, Jing Jiang 等ACM MM 2025 · 被引用 2 次
- Beyond Procedure: Substantive Fairness in Conformal PredictionPengqi Liu, Zijun Yu, Mouloud Belbahri, Arthur Charpentier 等ICML 2026 · 被引用 1 次
- Software Fairness Dilemma: Is Bias Mitigation a Zero-Sum Game?Zhenpeng Chen, Xinyue Li, Jie M. Zhang, Weisong Sun 等FSE 2025
- A Comprehensive Study of Deep Learning Model Fixing ApproachesHanmo You, Zan Wang, Zishuo Dong, Luanqi Mo 等ICSE 2026
- A Game-Theoretic Framework for Measuring and Explaining Metric Compatibility in Fair Machine LearningLingfeng Zhang, Jingran Yang, Zhaohui Wang, Min Zhang 等ICML 2026
它引用的顶会 Paper30
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 被引用 436 次
- Bias in machine learning software: why? how? what to do?Joymallya Chakraborty, Suvodeep Majumder, Tim MenziesFSE 2021 · 被引用 186 次
- Deep learning library testing via effective model generationZan Wang, Ming Yan, Junjie Chen, Shuang Liu 等FSE 2020 · 被引用 165 次
- FairBatch: Batch Selection for Model FairnessYuji Roh, Kangwook Lee, Steven Euijong Whang, Changho SuhICLR 2021 · 被引用 156 次
- Fairway: a way to build fair ML softwareJoymallya Chakraborty, Suvodeep Majumder, Zhe Yu, Tim MenziesFSE 2020 · 被引用 131 次
相关 Paper
- Adaptive fairness improvement based on causality analysisMengdi Zhang, Jun SunFSE 2022 · 被引用 35 次
- MEDFAIR: Benchmarking Fairness for Medical ImagingYongshuo Zong, Yongxin Yang, Timothy M. HospedalesICLR 2023 · 被引用 15 次
- Do the machine learning models on a crowd sourced platform exhibit bias? an empirical study on model fairnessSumon Biswas, Hridesh RajanFSE 2020 · 被引用 96 次
- Fair Without Leveling Down: A New Intersectional Fairness DefinitionGaurav Maheshwari, Aurélien Bellet, Pascal Denis, Mikaela KellerEMNLP 2023 · 被引用 3 次
- Are My Deep Learning Systems Fair? An Empirical Study of Fixed-Seed TrainingShangshu Qian, Hung Viet Pham, Thibaud Lutellier, Zeou Hu 等NeurIPS 2021 · 被引用 49 次
