Robust ML Auditing using Prior Knowledge
Jade Garcia Bourrée, Augustin Godinot, Sayan Biswas, Anne-Marie Kermarrec, Erwan Le Merrer, Gilles Trédan, Martijn de Vos, Milos Vujasinovic
Abstract
Among the many technical challenges to enforcing AI regulations, one crucial yet underexplored problem is the risk of audit manipulation. This manipulation occurs when a platform deliberately alters its answers to a regulator to pass an audit without modifying its answers to other users. In this paper, we introduce a novel approach to manipulation-proof auditing by taking into account the auditor's prior knowledge of the task solved by the platform. We first demonstrate that regulators must not rely on public priors (e.g., a public dataset), as platforms could easily fool the auditor in such cases. We then formally establish the conditions under which an auditor can prevent audit manipulations using prior knowledge about the ground truth. Finally, our experiments with two standard datasets illustrate the maximum level of unfairness a platform can hide before being detected as malicious. Our formalization and generalization of manipulation-proof auditing with a prior opens up new research directions for more robust fairness audits.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on20
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 671 citations
- Fairwashing explanations with off-manifold detergentChristopher J. Anders, Plamen Pasliev, Ann-Kathrin Dombrowski, Klaus-Robert Müller et al.ICML 2020 · 104 citations
- Too Relaxed to Be FairMichael Lohaus, Michaël Perrot, Ulrike von LuxburgICML 2020 · 80 citations
- Understanding Practices, Challenges, and Opportunities for User-Engaged Algorithm Auditing in Industry PracticeWesley Hanwen Deng, Bill Boyuan Guo, Alicia DeVrio, Hong Shen et al.CHI 2023 · 73 citations
- Sociotechnical Audits: Broadening the Algorithm Auditing Lens to Investigate Targeted AdvertisingMichelle S. Lam, Ayush Pandit, Colin H. Kalicki, Rachit Gupta et al.CSCW 2023 · 40 citations
Related papers
- Regulating algorithmic filtering on social mediaSarah Huiyi Cen, Devavrat ShahNeurIPS 2021 · 12 citations
- Having your Privacy Cake and Eating it Too: Platform-supported Auditing of Social Media Algorithms for Public InterestBasileal Imana, Aleksandra Korolova, John S. HeidemannCSCW 2023 · 18 citations
- 'I Know You Are Discriminatory!': Automated Substantiating for Individual Fairness Auditing of AI SystemsYuanhao Liu, Qi Cao, Huawei Shen, Kaike Zhang et al.CSCW 2025
- "Something Fast and Cheap" or "A Core Element of Building Trust"? - AI Auditing Professionals' Perspectives on Trust in AITina B. Lassiter, Kenneth R. FleischmannCSCW 2024 · 16 citations
- Beyond Bias Detection: Community Auditors and Normative Reasoning in AI Oversight CSCW006Corey Jackson, Tallal Ahmad, Shelcia David Raj, Natalie WuCSCW 2026 · 1 citation
