MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient Search
Zhaohui Wang, Min Zhang, Jingran Yang, Bojie Shao, Min Zhang
摘要
Deep neural networks (DNNs) have shown powerful performance in various applications and are increasingly being used in decisionmaking systems. However, concerns about fairness in DNNs always persist. Some efficient white-box fairness testing methods about individual fairness have been proposed. Nevertheless, the development of black-box methods has stagnated, and the performance of existing methods is far behind that of white-box methods. In this paper, we propose a novel black-box individual fairness testing method called Model-Agnostic Fairness Testing (MAFT). By leveraging MAFT, practitioners can effectively identify and address discrimination in DL models, regardless of the specific algorithm or architecture employed. Our approach adopts lightweight procedures such as gradient estimation and attribute perturbation rather than non-trivial procedures like symbol execution, rendering it significantly more scalable and applicable than existing methods. We demonstrate that MAFT achieves the same effectiveness as stateof-the-art white-box methods whilst improving the applicability to large-scale networks. Compared to existing black-box approaches, our approach demonstrates distinguished performance in discovering fairness violations w.r.t effectiveness (∼ 14.69×) and efficiency (∼ 32.58×). CCS CONCEPTS • Computing methodologies → Artificial intelligence; • Software and its engineering → Software testing and debugging.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language ModelsYisong Xiao, Aishan Liu, Siyuan Liang, Xianglong Liu 等ISSTA 2025 · 被引用 2 次
- Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness BugsRanit Debnath Akash, Ashish Kumar, Gang Tan, Saeid Tizpaz-NiariISSTA 2026
- Perturbation Effects on Robustness and Individual FairnessXuran Li, Hao Xue, Peng Wu, Xingjun Ma 等KDD 2026
- Provable Fairness Repair for Deep Neural NetworksJianan Ma, Jingyi Wang, Qi Xuan, Zhen WangASE 2025
它引用的顶会 Paper6
- Fairness-aware News Recommendation with Decomposed Adversarial LearningChuhan Wu, Fangzhao Wu, Xiting Wang, Yongfeng Huang 等AAAI 2021 · 被引用 176 次
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong 等ICSE 2020 · 被引用 127 次
- NeuronFair: Interpretable White-Box Fairness Testing through Biased Neuron IdentificationHaibin Zheng, Zhiqing Chen, Tianyu Du, Xuhong Zhang 等ICSE 2022 · 被引用 58 次
- Efficient white-box fairness testing through gradient searchLingfeng Zhang, Yueling Zhang, Min ZhangISSTA 2021 · 被引用 51 次
- Explanation-Guided Fairness Testing through Genetic AlgorithmMing Fan, Wenying Wei, Wuxia Jin, Zijiang Yang 等ICSE 2022 · 被引用 45 次
相关 Paper
- Approximation-guided Fairness Testing through Discriminatory Space AnalysisZhenjiang Zhao, Takahisa Toda, Takashi KitamuraASE 2024
- Dissecting Global Search: A Simple Yet Effective Method to Boost Individual Discrimination Testing and RepairLili Quan, Tianlin Li, Xiaofei Xie, Zhenpeng Chen 等ICSE 2025 · 被引用 2 次
- RULER: discriminative and iterative adversarial training for deep neural network fairnessGuanhong Tao, Weisong Sun, Tingxu Han, Chunrong Fang 等FSE 2022 · 被引用 29 次
- Fairquant: Certifying and Quantifying Fairness of Deep Neural NetworksBrian Hyeongseok Kim, Jingbo Wang, Chao WangICSE 2025 · 被引用 6 次
- Fairify: Fairness Verification of Neural NetworksSumon Biswas, Hridesh RajanICSE 2023 · 被引用 27 次
