Robust ML Auditing using Prior Knowledge
Jade Garcia Bourrée, Augustin Godinot, Sayan Biswas, Anne-Marie Kermarrec, Erwan Le Merrer, Gilles Trédan, Martijn de Vos, Milos Vujasinovic
摘要
Among the many technical challenges to enforcing AI regulations, one crucial yet underexplored problem is the risk of audit manipulation. This manipulation occurs when a platform deliberately alters its answers to a regulator to pass an audit without modifying its answers to other users. In this paper, we introduce a novel approach to manipulation-proof auditing by taking into account the auditor's prior knowledge of the task solved by the platform. We first demonstrate that regulators must not rely on public priors (e.g., a public dataset), as platforms could easily fool the auditor in such cases. We then formally establish the conditions under which an auditor can prevent audit manipulations using prior knowledge about the ground truth. Finally, our experiments with two standard datasets illustrate the maximum level of unfairness a platform can hide before being detected as malicious. Our formalization and generalization of manipulation-proof auditing with a prior opens up new research directions for more robust fairness audits.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 被引用 671 次
- Fairwashing explanations with off-manifold detergentChristopher J. Anders, Plamen Pasliev, Ann-Kathrin Dombrowski, Klaus-Robert Müller 等ICML 2020 · 被引用 104 次
- Too Relaxed to Be FairMichael Lohaus, Michaël Perrot, Ulrike von LuxburgICML 2020 · 被引用 80 次
- Understanding Practices, Challenges, and Opportunities for User-Engaged Algorithm Auditing in Industry PracticeWesley Hanwen Deng, Bill Boyuan Guo, Alicia DeVrio, Hong Shen 等CHI 2023 · 被引用 73 次
- Sociotechnical Audits: Broadening the Algorithm Auditing Lens to Investigate Targeted AdvertisingMichelle S. Lam, Ayush Pandit, Colin H. Kalicki, Rachit Gupta 等CSCW 2023 · 被引用 40 次
相关 Paper
- Regulating algorithmic filtering on social mediaSarah Huiyi Cen, Devavrat ShahNeurIPS 2021 · 被引用 12 次
- Having your Privacy Cake and Eating it Too: Platform-supported Auditing of Social Media Algorithms for Public InterestBasileal Imana, Aleksandra Korolova, John S. HeidemannCSCW 2023 · 被引用 18 次
- 'I Know You Are Discriminatory!': Automated Substantiating for Individual Fairness Auditing of AI SystemsYuanhao Liu, Qi Cao, Huawei Shen, Kaike Zhang 等CSCW 2025
- "Something Fast and Cheap" or "A Core Element of Building Trust"? - AI Auditing Professionals' Perspectives on Trust in AITina B. Lassiter, Kenneth R. FleischmannCSCW 2024 · 被引用 16 次
- Beyond Bias Detection: Community Auditors and Normative Reasoning in AI Oversight CSCW006Corey Jackson, Tallal Ahmad, Shelcia David Raj, Natalie WuCSCW 2026 · 被引用 1 次
