Correct-by-Construction: Certified Individual Fairness through Neural Network Training
Ruihan Zhang, Jun Sun
摘要
Fairness in machine learning is more important than ever as ethical concerns continue to grow. Individual fairness demands that individuals differing only in sensitive attributes receive the same outcomes. However, commonly used machine learning algorithms often fail to achieve such fairness. To improve individual fairness, various training methods have been developed, such as incorporating fairness constraints as optimisation objectives. While these methods have demonstrated empirical effectiveness, they lack formal guarantees of fairness. Existing approaches that aim to provide fairness guarantees primarily rely on verification techniques, which can sometimes fail to produce definitive results. Moreover, verification alone does not actively enhance individual fairness during training. To address this limitation, we propose a novel framework that formally guarantees individual fairness throughout training. Our approach consists of two parts, i.e., (1) provably fair initialisation that ensures the model starts in a fair state, and (2) a fairness-preserving training algorithm that maintains fairness as the model learns. A key element of our method is the use of randomised response mechanisms, which protect sensitive attributes while maintaining fairness guarantees. We formally prove that this mechanism sustains individual fairness throughout the training process. Experimental evaluations confirm that our approach is effective, i.e., producing models that are empirically fair and accurate. Furthermore, our approach is much more efficient than the alternative approach based on certified training (which requires neural network verification during training).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper20
- Formal Security Analysis of Neural Networks using Symbolic IntervalsShiqi Wang, Kexin Pei, Justin Whitehouse, Junfeng Yang 等USENIX Security 2018 · 被引用 523 次
- To be Robust or to be Fair: Towards Fairness in Adversarial TrainingHan Xu, Xiaorui Liu, Yaxin Li, Anil K. Jain 等ICML 2021 · 被引用 218 次
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong 等ICSE 2020 · 被引用 127 次
- Learning Certified Individually Fair RepresentationsAnian Ruoss, Mislav Balunovic, Marc Fischer, Martin T. VechevNeurIPS 2020 · 被引用 112 次
- An Abstraction-Based Framework for Neural Network VerificationYizhak Yisrael Elboher, Justin Gottschlich, Guy KatzCAV 2020 · 被引用 97 次
相关 Paper
- SenSeI: Sensitive Set Invariance for Enforcing Individual FairnessMikhail Yurochkin, Yuekai SunICLR 2021 · 被引用 53 次
- Training individually fair ML models with sensitive subspace robustnessMikhail Yurochkin, Amanda Bower, Yuekai SunICLR 2020 · 被引用 123 次
- CertiFair: A Framework for Certified Global Fairness of Neural NetworksHaitham Khedr, Yasser ShoukryAAAI 2023 · 被引用 26 次
- Certification of Distributional Individual FairnessMatthew Wicker, Vihari Piratla, Adrian WellerNeurIPS 2023 · 被引用 2 次
- Fairify: Fairness Verification of Neural NetworksSumon Biswas, Hridesh RajanICSE 2023 · 被引用 27 次
