Causal Context Connects Counterfactual Fairness to Robust Prediction and Group Fairness
Jacy Reese Anthis, Victor Veitch
Abstract
Counterfactual fairness requires that a person would have been classified in the same way by an AI or other algorithmic system if they had a different protected class, such as a different race or gender. This is an intuitive standard, as reflected in the U.S. legal system, but its use is limited because counterfactuals cannot be directly observed in real-world data. On the other hand, group fairness metrics (e.g., demographic parity or equalized odds) are less intuitive but more readily observed. In this paper, we use to bridge the gaps between counterfactual fairness, robust prediction, and group fairness. First, we motivate counterfactual fairness by showing that there is not necessarily a fundamental trade-off between fairness and accuracy because, under plausible conditions, the counterfactually fair predictor is in fact accuracy-optimal in an unbiased target distribution. Second, we develop a correspondence between the causal graph of the data-generating process and which, if any, group fairness metrics are equivalent to counterfactual fairness. Third, we show that in three common fairness contextsmeasurement error, selection on label, and selection on predictorscounterfactual fairness is equivalent to demographic parity, equalized odds, and calibration, respectively. Counterfactual fairness can sometimes be tested by measuring relatively simple group fairness metrics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c5b423e9-14d8-45c7-8e45-69b863ae7b75Cited by top-tier papers8
- Bias in Language Models: Beyond Trick Tests and Towards RUTEd EvaluationKristian Lum, Jacy Reese Anthis, Kevin Robinson, Chirag Nagpal et al.ACL 2025 · 41 citations
- Counterfactual Fairness by Combining Factual and Counterfactual PredictionsZeyu Zhou, Tianci Liu, Ruqi Bai, Jing Gao et al.NeurIPS 2024 · 11 citations
- Mind the Graph When Balancing Data for Fairness or RobustnessJessica Schrouff, Alexis Bellot, Amal Rannen-Triki, Alan Malek et al.NeurIPS 2024 · 10 citations
- Learning Counterfactual Outcomes Under Rank PreservationPeng Wu, Haoxuan Li, Chunyuan Zheng, Yan Zeng et al.NeurIPS 2025 · 7 citations
- Certifying Counterfactual Bias in LLMsIsha Chaudhary, Qian Hu, Manoj Kumar, Morteza Ziyadi et al.ICLR 2025 · 3 citations
Builds on13
- On Calibration and Out-of-Domain GeneralizationYoav Wald, Amir Feder, Daniel Greenfeld, Uri ShalitNeurIPS 2021 · 184 citations
- Is There a Trade-Off Between Fairness and Accuracy? A Perspective Using Mismatched Hypothesis TestingSanghamitra Dutta, Dennis Wei, Hazar Yueksel, Pin-Yu Chen et al.ICML 2020 · 171 citations
- Conditional Learning of Fair RepresentationsHan Zhao, Amanda Coston, Tameem Adel, Geoffrey J. GordonICLR 2020 · 127 citations
- Counterfactual Invariance to Spurious Correlations in Text ClassificationVictor Veitch, Alexander D'Amour, Steve Yadlowsky, Jacob EisensteinNeurIPS 2021 · 108 citations
- Achieving Fairness at No Utility Cost via Data Reweighing with InfluencePeizhao Li, Hongfu LiuICML 2022 · 57 citations
Related papers
- On the Fairness of Causal Algorithmic RecourseJulius von Kügelgen, Amir-Hossein Karimi, Umang Bhatt, Isabel Valera et al.AAAI 2022 · 99 citations
- Causal Conceptions of Fairness and their ConsequencesHamed Nilforoshan, Johann D. Gaebler, Ravi Shroff, Sharad GoelICML 2022 · 52 citations
- Learning for Counterfactual Fairness from Observational DataJing Ma, Ruocheng Guo, Aidong Zhang, Jundong LiKDD 2023 · 9 citations
- A Causal Look at Statistical Definitions of DiscriminationElias Chaibub NetoKDD 2020 · 3 citations
- Pairwise Fairness for Ranking and RegressionHarikrishna Narasimhan, Andrew Cotter, Maya R. Gupta, Serena Lutong WangAAAI 2020 · 125 citations
