Constructing a Fair Classifier with Generated Fair Data
Taeuk Jang, Feng Zheng, Xiaoqian Wang
Abstract
Fairness in machine learning is getting rising attention as it is directly related to real-world applications and social problems. Recent methods have been explored to alleviate the discrimination between certain demographic groups that are characterized by sensitive attributes (such as race, age, or gender). Some studies have found that the data itself is biased, so training directly on the data causes unfair decision making. Models directly trained on raw data can replicate or even exacerbate bias in the prediction between demographic groups. This leads to vastly different prediction performance in different demographic groups. In order to address this issue, we propose a new approach to improve machine learning fairness by generating fair data. We introduce a generative model to generate cross-domain samples w.r.t. multiple sensitive attributes. This ensures that we can generate infinite number of samples that are balanced w.r.t. both target label and sensitive attributes to enhance fair prediction. By training the classifier solely with the synthetic data and then transfer the model to real data, we can overcome the under-representation problem which is non-trivial since collecting real data is extremely time and resource consuming. We provide empirical evidence to demonstrate the benefit of our model with respect to both fairness and accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2d995a06-eb8b-4f84-b93b-dbf26e460e89Cited by top-tier papers12
- Fairness without Demographics through Knowledge DistillationJunyi Chai, Taeuk Jang, Xiaoqian WangNeurIPS 2022 · 57 citations
- Fairness with Adaptive WeightsJunyi Chai, Xiaoqian WangICML 2022 · 47 citations
- Self-Supervised Fair Representation Learning without DemographicsJunyi Chai, Xiaoqian WangNeurIPS 2022 · 35 citations
- Causal Context Connects Counterfactual Fairness to Robust Prediction and Group FairnessJacy Reese Anthis, Victor VeitchNeurIPS 2023 · 26 citations
- A Fair Generative Model Using LeCam DivergenceSoobin Um, Changho SuhAAAI 2023 · 8 citations
Related papers
- Towards Accuracy-Fairness Paradox: Adversarial Example-based Data Augmentation for Visual DebiasingYi Zhang, Jitao SangACM MM 2020 · 32 citations
- Robust Fairness Under Covariate ShiftAshkan Rezaei, Anqi Liu, Omid Memarrast, Brian D. ZiebartAAAI 2021 · 94 citations
- Gradient Based Activations for Accurate Bias-Free LearningVinod K. Kurmi, Rishabh Sharma, Yash Vardhan Sharma, Vinay P. NamboodiriAAAI 2022 · 3 citations
- Graph Fairness Learning under Distribution ShiftsYibo Li, Xiao Wang, Yujie Xing, Shaohua Fan et al.WWW 2024 · 16 citations
- Learning for Counterfactual Fairness from Observational DataJing Ma, Ruocheng Guo, Aidong Zhang, Jundong LiKDD 2023 · 9 citations
