Fair Wrapping for Black-box Predictions
Alexander Soen, Ibrahim M. Alabdulmohsin, Sanmi Koyejo, Yishay Mansour, Nyalleng Moorosi, Richard Nock, Ke Sun, Lexing Xie
Abstract
We introduce a new family of techniques to post-process ("wrap") a black-box classifier in order to reduce its bias. Our technique builds on the recent analysis of improper loss functions whose optimization can correct any twist in prediction, unfairness being treated as a twist. In the post-processing, we learn a wrapper function which we define as an -tree, which modifies the prediction. We provide two generic boosting algorithms to learn -trees. We show that our modification has appealing properties in terms of composition of -trees, generalization, interpretability, and KL divergence between modified and original predictions. We exemplify the use of our technique in three fairness notions: conditional value-at-risk, equality of opportunity, and statistical parity; and provide experiments on several readily available datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 720591a9-83dd-447f-8b7f-b7505529bd8eCited by top-tier papers1
Ask how each one uses itBuilds on6
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 671 citations
- Post-processing for Individual FairnessFelix Petersen, Debarghya Mukherjee, Yuekai Sun, Mikhail YurochkinNeurIPS 2021 · 115 citations
- Fairness with Overlapping Groups; a Probabilistic PerspectiveForest Yang, Mouhamadou Cisse, Oluwasanmi KoyejoNeurIPS 2020 · 71 citations
- Use Privacy in Data-Driven Systems: Theory and Experiments with Machine Learnt ProgramsAnupam Datta, Matthew Fredrikson, Gihyuk Ko, Piotr Mardziel et al.CCS 2017 · 63 citations
- A Near-Optimal Algorithm for Debiasing Trained Machine Learning ModelsIbrahim M. Alabdulmohsin, Mario LucicNeurIPS 2021 · 26 citations
Related papers
- Boosted CVaR ClassificationRuntian Zhai, Chen Dan, Arun Sai Suggala, J. Zico Kolter et al.NeurIPS 2021 · 17 citations
- A Causal Look at Statistical Definitions of DiscriminationElias Chaibub NetoKDD 2020 · 3 citations
- Being Properly ImproperTyler Sypherd, Richard Nock, Lalitha SankarICML 2022 · 14 citations
- Post-hoc bias scoring is optimal for fair classificationWenlong Chen, Yegor Klochkov, Yang LiuICLR 2024 · 12 citations
- FRAPPÉ: A Group Fairness Framework for Post-Processing EverythingAlexandru Tifrea, Preethi Lahoti, Ben Packer, Yoni Halpern et al.ICML 2024 · 15 citations
