Fairquant: Certifying and Quantifying Fairness of Deep Neural Networks
Brian Hyeongseok Kim, Jingbo Wang, Chao Wang
Abstract
We propose a method for formally certifying and quantifying individual fairness of deep neural networks (DNN). Individual fairness guarantees that any two individuals who are identical except for a legally protected attribute (e.g., gender or race) receive the same treatment. While there are existing techniques that provide such a guarantee, they tend to suffer from lack of scalability or accuracy as the size and input dimension of the DNN increase. Our method overcomes this limitation by applying abstraction to a symbolic interval based analysis of the DNN followed by iterative refinement guided by the fairness property. Furthermore, our method lifts the symbolic interval based analysis from conventional qualitative certification to quantitative certification, by computing the percentage of individuals whose classification outputs are provably fair, instead of merely deciding if the DNN is fair. We have implemented our method and evaluated it on deep neural networks trained on four popular fairness research datasets. The experimental results show that our method is not only more accurate than state-of-the-art techniques but also several orders-of-magnitude faster. Certification Problem ⟨f, xj , X⟩ ∃P ∈ stack S to certify? Fair, Unfair, and Undecided rates (%) initial partition P ← X added to S No Yes Certification Subproblem ⟨f, xj , P ⟩ Abstraction (Forward Analysis) Refinement (Backward Analysis) Quantification (Rate Computation)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Dissecting Global Search: A Simple Yet Effective Method to Boost Individual Discrimination Testing and RepairLili Quan, Tianlin Li, Xiaofei Xie, Zhenpeng Chen et al.ICSE 2025 · 2 citations
- Correct-by-Construction: Certified Individual Fairness through Neural Network TrainingRuihan Zhang, Jun SunOOPSLA 2025 · 1 citation
- Quantifying Sensitivity for Tree Ensembles: A Symbolic and Compositional ApproachAjinkya Naik, Chaitanya Garg, S. Akshay, Ashutosh Gupta et al.CAV 2026
- Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness BugsRanit Debnath Akash, Ashish Kumar, Gang Tan, Saeid Tizpaz-NiariISSTA 2026
- Solving Probabilistic Verification Problems of Neural Networks using Branch and BoundDavid Boetius, Stefan Leue, Tobias SutterICML 2025
Builds on15
- Formal Security Analysis of Neural Networks using Symbolic IntervalsShiqi Wang, Kexin Pei, Justin Whitehouse, Junfeng Yang et al.USENIX Security 2018 · 523 citations
- Fairness without Demographics through Adversarially Reweighted LearningPreethi Lahoti, Alex Beutel, Jilin Chen, Kang Lee et al.NeurIPS 2020 · 406 citations
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong et al.ICSE 2020 · 127 citations
- Training individually fair ML models with sensitive subspace robustnessMikhail Yurochkin, Amanda Bower, Yuekai SunICLR 2020 · 123 citations
- Learning Certified Individually Fair RepresentationsAnian Ruoss, Mislav Balunovic, Marc Fischer, Martin T. VechevNeurIPS 2020 · 112 citations
Related papers
- Fairify: Fairness Verification of Neural NetworksSumon Biswas, Hridesh RajanICSE 2023 · 27 citations
- Certification of Distributional Individual FairnessMatthew Wicker, Vihari Piratla, Adrian WellerNeurIPS 2023 · 2 citations
- CertiFair: A Framework for Certified Global Fairness of Neural NetworksHaitham Khedr, Yasser ShoukryAAAI 2023 · 26 citations
- Provable Fairness Repair for Deep Neural NetworksJianan Ma, Jingyi Wang, Qi Xuan, Zhen WangASE 2025
- Azkaban: A Zero-Knowledge Abstract Analysis for Neural NetworksSankha Das, Lucien K. L. Ng, Yibin Yang, Vladimir Kolesnikov et al.CCS 2026
