FARE: Provably Fair Representation Learning with Practical Certificates
Nikola Jovanovic, Mislav Balunovic, Dimitar Iliev Dimitrov, Martin T. Vechev
Abstract
Fair representation learning (FRL) is a popular class of methods aiming to produce fair classifiers via data preprocessing. Recent regulatory directives stress the need for FRL methods that provide practical certificates, i.e., provable upper bounds on the unfairness of any downstream classifier trained on preprocessed data, which directly provides assurance in a practical scenario. Creating such FRL methods is an important challenge that remains unsolved. In this work, we address that challenge and introduce FARE (Fairness with Restricted Encoders), the first FRL method with practical fairness certificates. FARE is based on our key insight that restricting the representation space of the encoder enables the derivation of practical guarantees, while still permitting favorable accuracy-fairness tradeoffs for suitable instantiations, such as one we propose based on fair trees. To produce a practical certificate, we develop and apply a statistical procedure that computes a finite sample high-confidence upper bound on the unfairness of any downstream classifier trained on FARE embeddings. In our comprehensive experimental evaluation, we demonstrate that FARE produces practical certificates that are tight and often even comparable with purely empirical results obtained by prior methods, which establishes the practical value of our approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 208cebba-720d-48b2-8223-d21729cb2b4cCited by top-tier papers9
- Endowing Pre-trained Graph Models with Provable FairnessZhongjian Zhang, Mengmei Zhang, Yue Yu, Cheng Yang et al.WWW 2024 · 16 citations
- From Principle to Practice: Vertical Data Minimization for Machine LearningRobin Staab, Nikola Jovanovic, Mislav Balunovic, Martin T. VechevS&P 2024 · 10 citations
- FedFACT: A Provable Framework for Controllable Group-Fairness Calibration in Federated LearningLi Zhang, Zhongxuan Han, Xiaohua Feng, Jiaming Zhang et al.NeurIPS 2025 · 2 citations
- Fairness-aware Anomaly Detection via Fair ProjectionFeng Xiao, Xiaoying Tang, Jicong FanNeurIPS 2025 · 2 citations
- Efficient Fairness-Performance Pareto Front ComputationMark Kozdoba, Binyamin Perets, Shie MannorNeurIPS 2025 · 2 citations
Builds on19
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan et al.CCS 2016 · 7,620 citations
- AI2: Safety and Robustness Certification of Neural Networks with Abstract InterpretationTimon Gehr, Matthew Mirman, Dana Drachsler-Cohen, Petar Tsankov et al.S&P 2018 · 987 citations
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 671 citations
- A Theory of Usable Information under Computational ConstraintsYilun Xu, Shengjia Zhao, Jiaming Song, Russell Stewart et al.ICLR 2020 · 211 citations
- Conditional Learning of Fair RepresentationsHan Zhao, Amanda Coston, Tameem Adel, Geoffrey J. GordonICLR 2020 · 127 citations
Related papers
- Learning Certified Individually Fair RepresentationsAnian Ruoss, Mislav Balunovic, Marc Fischer, Martin T. VechevNeurIPS 2020 · 112 citations
- Fair Representation Learning with Controllable High Confidence Guarantees via Adversarial InferenceYuhong Luo, Austin Hoag, Xintong Wang, Philip S. Thomas et al.NeurIPS 2025
- Utility-Fairness Trade-Offs and how to Find ThemSepehr Dehdashtian, Bashir Sadeghi, Vishnu Naresh BoddetiCVPR 2024
- Fair Normalizing FlowsMislav Balunovic, Anian Ruoss, Martin T. VechevICLR 2022 · 46 citations
- Controllable Universal Fair Representation LearningYue Cui, Chen Ma, Kai Zheng, Lei Chen et al.WWW 2023 · 5 citations
