Probabilistic Verification of Neural Networks Against Group Fairness
Bing Sun, Jun Sun, Ting Dai, Lijun Zhang
摘要
Fairness is crucial for neural networks which are used in applications with important societal implication. Recently, there have been multiple attempts on improving fairness of neural networks, with a focus on fairness testing (e.g., generating individual discriminatory instances) and fairness training (e.g., enhancing fairness through augmented training). In this work, we propose an approach to formally verify neural networks against fairness, with a focus on independence-based fairness such as group fairness. Our method is built upon an approach for learning Markov Chains from a user-provided neural network (i.e., a feed-forward neural network or a recurrent neural network) which is guaranteed to facilitate sound analysis. The learned Markov Chain not only allows us to verify (with Probably Approximate Correctness guarantee) whether the neural network is fair or not, but also facilities sensitivity analysis which helps to understand why fairness is violated. We demonstrate that with our analysis results, the neural weights can be optimized to improve fairness. Our approach has been evaluated with multiple models trained on benchmark datasets and the experiment results show that our approach is effective and efficient.
We have implemented our approach as a part of the SOCRATES framework [45]. We apply our approach to multiple neural network models (including feed-forward and recurrent neural networks) trained on benchmark datasets which are the subject of previous studies on fairness testing. The experiment results show that our approach successfully verifies or falsifies all the models. It also confirms that fairness is a real concern and one of the networks (on the
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Causality-Based Neural Network RepairBing Sun, Jun Sun, Long H. Pham, Tie ShiICSE 2022 · 被引用 69 次
- Monitoring Algorithmic FairnessThomas A. Henzinger, Mahyar Karimi, Konstantin Kueffner, Kaushik MallikCAV 2023 · 被引用 13 次
- Semantic-Based Neural Network RepairRichard Schumi, Jun SunISSTA 2023 · 被引用 7 次
- Certified Continual Learning for Neural Network RegressionLong H. Pham, Jun SunISSTA 2024 · 被引用 2 次
- Fairness Shields: Safeguarding against Biased Decision MakersFilip Cano, Thomas A. Henzinger, Bettina Könighofer, Konstantin Kueffner 等AAAI 2025
它引用的顶会 Paper6
- AI2: Safety and Robustness Certification of Neural Networks with Abstract InterpretationTimon Gehr, Matthew Mirman, Dana Drachsler-Cohen, Petar Tsankov 等S&P 2018 · 被引用 987 次
- TextBugger: Generating Adversarial Text Against Real-world ApplicationsJinfeng Li, Shouling Ji, Tianyu Du, Bo Li 等NDSS 2019 · 被引用 876 次
- Formal Security Analysis of Neural Networks using Symbolic IntervalsShiqi Wang, Kexin Pei, Justin Whitehouse, Junfeng Yang 等USENIX Security 2018 · 被引用 523 次
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong 等ICSE 2020 · 被引用 127 次
- An Abstraction-Based Framework for Neural Network VerificationYizhak Yisrael Elboher, Justin Gottschlich, Guy KatzCAV 2020 · 被引用 97 次
相关 Paper
- Fairify: Fairness Verification of Neural NetworksSumon Biswas, Hridesh RajanICSE 2023 · 被引用 27 次
- Adaptive fairness improvement based on causality analysisMengdi Zhang, Jun SunFSE 2022 · 被引用 35 次
- Fairquant: Certifying and Quantifying Fairness of Deep Neural NetworksBrian Hyeongseok Kim, Jingbo Wang, Chao WangICSE 2025 · 被引用 6 次
- CertiFair: A Framework for Certified Global Fairness of Neural NetworksHaitham Khedr, Yasser ShoukryAAAI 2023 · 被引用 26 次
- Correct-by-Construction: Certified Individual Fairness through Neural Network TrainingRuihan Zhang, Jun SunOOPSLA 2025 · 被引用 1 次
