Active fairness auditing
Tom Yan, Chicheng Zhang
Abstract
The fast spreading adoption of machine learning (ML) by companies across industries poses significant regulatory challenges. One such challenge is scalability: how can regulatory bodies efficiently audit these ML models, ensuring that they are fair? In this paper, we initiate the study of query-based auditing algorithms that can estimate the demographic parity of ML models in a query-efficient manner. We propose an optimal deterministic algorithm, as well as a practical randomized, oracle-efficient algorithm with comparable guarantees. Furthermore, we make inroads into understanding the optimal query complexity of randomized active fairness estimation algorithms. Our first exploration of active fairness estimation aims to put AI governance on firmer theoretical foundations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic VariationsDavid Hartmann, Amin Oueslati, Dimitri Staufer, Lena Pohlmann et al.CHI 2025 · 37 citations
- FairProof : Confidential and Certifiable Fairness for Neural NetworksChhavi Yadav, Amrita Roy Chowdhury, Dan Boneh, Kamalika ChaudhuriICML 2024 · 20 citations
- Log Probability Tracking of LLM APIsTimothee Chauvin, Erwan Le Merrer, Francois Taiani, Gilles TredanICLR 2026 · 12 citations
- Demystifying Local & Global Fairness Trade-offs in Federated Learning Using Partial Information DecompositionFaisal Hamman, Sanghamitra DuttaICLR 2024 · 9 citations
- Cross-GAN Auditing: Unsupervised Identification of Attribute Level Similarities and Differences Between Pretrained Generative ModelsMatthew L. Olson, Shusen Liu, Rushil Anirudh, Jayaraman J. Thiagarajan et al.CVPR 2023
Builds on3
- Auditing Black-Box Prediction Models for Data Minimization ComplianceBashir Rastegarpanah, Krishna P. Gummadi, Mark CrovellaNeurIPS 2021 · 24 citations
- Estimating decision tree learnability with polylogarithmic sample complexityGuy Blanc, Neha Gupta, Jane Lange, Li-Yang TanNeurIPS 2020 · 5 citations
- VC dimension and distribution-free sample-based testingEric Blais, Renato Ferreira Pinto Jr., Nathaniel HarmsSTOC 2021
Related papers
- Active Fourier Auditor for Estimating Distributional Properties of ML ModelsAyoub Ajarra, Bishwamittra Ghosh, Debabrota BasuAAAI 2025 · 5 citations
- Audits Under Resource, Data, and Access Constraints: Scaling Laws For Less Discriminatory AlternativesSarah H. Cen, Salil Goyal, Zaynah Javed, Ananya Karthik et al.NeurIPS 2025 · 3 citations
- Stochastic Differentially Private and Fair LearningAndrew Lowy, Devansh Gupta, Meisam RazaviyaynICLR 2023 · 1 citation
- Meta Optimality for Demographic Parity Constrained Regression via Post-ProcessingKazuto FukuchiICML 2025
- Fair regression via plug-in estimator and recalibration with statistical guaranteesEvgenii Chzhen, Christophe Denis, Mohamed Hebiri, Luca Oneto et al.NeurIPS 2020 · 52 citations
