'I Know You Are Discriminatory!': Automated Substantiating for Individual Fairness Auditing of AI Systems
Yuanhao Liu, Qi Cao, Huawei Shen, Kaike Zhang, Yunfan Wu, Xueqi Cheng
摘要
Artificial intelligence (AI) systems are playing an increasingly crucial role in people's lives, and their frequent unfair behaviors raise concerns about fairness. To unveil the unfairness in AI systems, researchers conduct fairness auditing on these systems. However, existing fairness auditing works often focus on group fairness while ignoring discriminatory phenomena among individuals. To unearth discriminatory phenomena against individuals within AI systems, this paper proposes an individual fairness auditing framework, termed ''substantiating'', which can identify discrimination instances within AI systems by constructing individual samples. To construct these samples for substantiating, auditors often have to rely on subjective prior knowledge, lacking guidelines on how to construct unfair samples. To address this issue, this paper introduces two categories of automated sample generation methods that can rapidly find unfair samples within a limited number of queries to the system and demonstrate their effectiveness through experiments. This paper evaluates the proposed auditing framework among three categories of stakeholders in AI fairness: auditors, AI model developers, and non-technical personnel. The research findings point out their demand for individual fairness audits of AI systems and highlight how the framework supports a reliable and convenient individual fairness audit.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Efficient white-box fairness testing through gradient searchLingfeng Zhang, Yueling Zhang, Min ZhangISSTA 2021 · 被引用 51 次
- Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness TestingYisong Xiao, Aishan Liu, Tianlin Li, Xianglong LiuISSTA 2023 · 被引用 31 次
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong 等ICSE 2020 · 被引用 127 次
- Approximation-guided Fairness Testing through Discriminatory Space AnalysisZhenjiang Zhao, Takahisa Toda, Takashi KitamuraASE 2024
- WeAudit: Scaffolding User Auditors and AI Practitioners in Auditing Generative AIWesley Hanwen Deng, Claire Wang, Howard Ziyu Han, Jason I. Hong 等CSCW 2025 · 被引用 12 次
