FairPFN: A Tabular Foundation Model for Causal Fairness
Jake Robertson, Noah Hollmann, Samuel Müller, Noor H. Awad, Frank Hutter
Abstract
Machine learning (ML) systems are utilized in critical sectors, such as healthcare, law enforcement, and finance. However, these systems are often trained on historical data that contains demographic biases, leading to ML decisions that perpetuate or exacerbate existing social inequalities. Causal fairness provides a transparent, human-inthe-loop framework to mitigate algorithmic discrimination, aligning closely with legal doctrines of direct and indirect discrimination. However, current causal fairness frameworks hold a key limitation in that they assume prior knowledge of the correct causal model, restricting their applicability in complex fairness scenarios where causal models are unknown or difficult to identify. To bridge this gap, we propose FairPFN, a tabular foundation model pre-trained on synthetic causal fairness data to identify and mitigate the causal effects of protected attributes in its predictions. FairPFN's key contribution is that it requires no knowledge of the causal model and still demonstrates strong performance in identifying and removing protected causal effects across a diverse set of hand-crafted and real-world scenarios relative to robust baseline methods. FairPFN paves the way for promising future research, making causal fairness more accessible to a wider variety of complex fairness problems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7c8ad669-76d3-440b-8500-2772beb8a5f4Cited by top-tier papers3
- From Zero to Hero: Advancing Zero-Shot Foundation Models for Tabular Outlier DetectionXueying Ding, Haomin Wen, Simon Klüttermann, Leman AkogluICML 2026 · 5 citations
- pTNAS: Progressive Neural Architecture Search for Tabular DataNaili Xing, Shaofeng Cai, Lingze Zeng, Jiaqi Zhu et al.ICML 2026 · 4 citations
- Fair Data Pre-Processing with Imperfect Attribute SpaceYing Zheng, Yangfan Jiang, Kian-Lee TanSIGMOD 2026
Builds on3
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 671 citations
- Transformers Can Do Bayesian InferenceSamuel Müller, Noah Hollmann, Sebastian Pineda-Arango, Josif Grabocka et al.ICLR 2022 · 287 citations
- Learning for Counterfactual Fairness from Observational DataJing Ma, Ruocheng Guo, Aidong Zhang, Jundong LiKDD 2023 · 9 citations
Related papers
- Interventional Fairness on Partially Known Causal Graphs: A Constrained Optimization ApproachAoqi Zuo, Yiqing Li, Susan Wei, Mingming GongICLR 2024 · 10 citations
- Counterfactual Fairness with Partially Known Causal GraphAoqi Zuo, Susan Wei, Tongliang Liu, Bo Han et al.NeurIPS 2022 · 32 citations
- DECAF: Generating Fair Synthetic Data Using Causally-Aware Generative NetworksBoris van Breugel, Trent Kyono, Jeroen Berrevoets, Mihaela van der SchaarNeurIPS 2021 · 174 citations
- CausalPre: Scalable and Effective Data Pre-Processing for Causal FairnessYing Zheng, Yangfan Jiang, Kian-Lee TanICDE 2026
- Counterfactual Fairness by Combining Factual and Counterfactual PredictionsZeyu Zhou, Tianci Liu, Ruqi Bai, Jing Gao et al.NeurIPS 2024 · 11 citations
