RULER: discriminative and iterative adversarial training for deep neural network fairness
Guanhong Tao, Weisong Sun, Tingxu Han, Chunrong Fang, Xiangyu Zhang
Abstract
Deep Neural Networks (DNNs) are becoming an integral part of many real-world applications, such as autonomous driving and financial management. While these models enable autonomy, there are however concerns regarding their ethics in decision making. For instance, fairness is an aspect that requires particular attention. A number of fairness testing techniques have been proposed to address this issue, e.g., by generating test cases called individual discriminatory instances for repairing DNNs. Although they have demonstrated great potential, they tend to generate many test cases that are not directly effective in improving fairness and incur substantial computation overhead. We propose a new model repair technique, Ruler, by discriminating sensitive and non-sensitive attributes during test case generation for model repair. The generated cases are then used in training to improve DNN fairness. Ruler balances the trade-off between accuracy and fairness by decomposing the training procedure into two phases and introducing a novel iterative adversarial training method for fairness. Compared to the state-of-the-art techniques on four datasets, Ruler has 7-28 times more effective repair test cases generated, is 10-15 times faster in test generation, and has 26-43% more fairness improvement on average.
• Software and its engineering → Software testing and debugging; • Computing methodologies → Neural networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5cc47511-8338-4eca-a4b2-b0662ce00cc3Cited by top-tier papers8
- Fairness Improvement with Multiple Protected Attributes: How Far Are We?Zhenpeng Chen, Jie M. Zhang, Federica Sarro, Mark HarmanICSE 2024 · 33 citations
- Fix Fairness, Don't Ruin Accuracy: Performance Aware Fairness Repair using AutoMLGiang Nguyen, Sumon Biswas, Hridesh RajanFSE 2023 · 15 citations
- REGLO: Provable Neural Network Repair for Global Robustness PropertiesFeisi Fu, Zhilu Wang, Weichao Zhou, Yixuan Wang et al.AAAI 2024 · 11 citations
- NeuFair: Neural Network Fairness Repair with DropoutVishnu Asutosh Dasu, Ashish Kumar, Saeid Tizpaz-Niari, Gang TanISSTA 2024 · 9 citations
- Causality-Aided Trade-Off Analysis for Machine Learning FairnessZhenlan Ji, Pingchuan Ma, Shuai Wang, Yanhui LiASE 2023 · 6 citations
Builds on13
- Fast is better than free: Revisiting adversarial trainingEric Wong, Leslie Rice, J. Zico KolterICLR 2020 · 1,352 citations
- HopSkipJumpAttack: A Query-Efficient Decision-Based AttackJianbo Chen, Michael I. Jordan, Martin J. WainwrightS&P 2020 · 797 citations
- Nesterov Accelerated Gradient and Scale Invariance for Adversarial AttacksJiadong Lin, Chuanbiao Song, Kun He, Liwei Wang et al.ICLR 2020 · 765 citations
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong et al.ICSE 2020 · 127 citations
- Heuristic Black-Box Adversarial Attacks on Video Recognition ModelsZhipeng Wei, Jingjing Chen, Xingxing Wei, Linxi Jiang et al.AAAI 2020 · 84 citations
Related papers
- Efficient white-box fairness testing through gradient searchLingfeng Zhang, Yueling Zhang, Min ZhangISSTA 2021 · 51 citations
- NeuronFair: Interpretable White-Box Fairness Testing through Biased Neuron IdentificationHaibin Zheng, Zhiqing Chen, Tianyu Du, Xuhong Zhang et al.ICSE 2022 · 58 citations
- RUNNER: Responsible UNfair NEuron Repair for Enhancing Deep Neural Network FairnessTianlin Li, Yue Cao, Jian Zhang, Shiqian Zhao et al.ICSE 2024 · 11 citations
- Provable Fairness Repair for Deep Neural NetworksJianan Ma, Jingyi Wang, Qi Xuan, Zhen WangASE 2025
- Fairneuron: Improving Deep Neural Network Fairness with Adversary Games on Selective NeuronsXuanqi Gao, Juan Zhai, Shiqing Ma, Chao Shen et al.ICSE 2022 · 36 citations
