Audits Under Resource, Data, and Access Constraints: Scaling Laws For Less Discriminatory Alternatives
Sarah H. Cen, Salil Goyal, Zaynah Javed, Ananya Karthik, Percy Liang, Daniel E. Ho
Abstract
AI audits play a critical role in AI accountability and safety. One branch of the law for which AI audits are particularly salient is anti-discrimination law. Several areas of anti-discrimination law implicate the"less discriminatory alternative"(LDA) requirement, in which a protocol (e.g., model) is defensible if no less discriminatory protocol that achieves comparable performance can be found with a reasonable amount of effort. Notably, the burden of proving an LDA exists typically falls on the claimant (the party alleging discrimination). This creates a significant hurdle in AI cases, as the claimant would seemingly need to train a less discriminatory yet high-performing model, a task requiring resources and expertise beyond most litigants. Moreover, developers often shield information about and access to their model and training data as trade secrets, making it difficult to reproduce a similar model from scratch. In this work, we present a procedure enabling claimants to determine if an LDA exists, even when they have limited compute, data, information, and model access. We focus on the setting in which fairness is given by demographic parity and performance by binary cross-entropy loss. As our main result, we provide a novel closed-form upper bound for the loss-fairness Pareto frontier (PF). We show how the claimant can use it to fit a PF in the"low-resource regime,"then extrapolate the PF that applies to the (large) model being contested, all without training a single large model. The expression thus serves as a scaling law for loss-fairness PFs. To use this scaling law, the claimant would require a small subsample of the train/test data. Then, the claimant can fit the context-specific PF by training as few as 7 (small) models. We stress test our main result in simulations, finding that our scaling law holds even when the exact conditions of our theory do not.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4c71b0e7-1d8e-4aca-8f44-f50cde989668Cited by top-tier papers1
Ask how each one uses itBuilds on6
- Learning the Pareto Front with HypernetworksAviv Navon, Aviv Shamsian, Ethan Fetaya, Gal ChechikICLR 2021 · 189 citations
- FACT: A Diagnostic for Group Fairness Trade-offsJoon Sik Kim, Jiahao Chen, Ameet TalwalkarICML 2020 · 67 citations
- Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit ToolingVictor Ojewale, Ryan Steed, Briana Vecchione, Abeba Birhane et al.CHI 2025 · 46 citations
- Sociotechnical Audits: Broadening the Algorithm Auditing Lens to Investigate Targeted AdvertisingMichelle S. Lam, Ayush Pandit, Colin H. Kalicki, Rachit Gupta et al.CSCW 2023 · 40 citations
- Trustless Audits without Revealing Data or ModelsSuppakit Waiwitlikhit, Ion Stoica, Yi Sun, Tatsunori Hashimoto et al.ICML 2024 · 20 citations
Related papers
- Active fairness auditingTom Yan, Chicheng ZhangICML 2022 · 34 citations
- Confidential-PROFITT: Confidential PROof of FaIr Training of TreesAli Shahin Shamsabadi, Sierra Calanda Wyllie, Nicholas Franzese, Natalie Dullerud et al.ICLR 2023
- Estimating and Controlling for Equalized Odds via Sensitive Attribute PredictorsBeepul Bharti, Paul H. Yi, Jeremias SulamNeurIPS 2023 · 8 citations
- RECAST: Model Reconstruction via Counterfactual-Aware Wasserstein Geometry under Limited DataXuan Zhao, Lena Krieger, Zhuo Cao, Arya Bangun et al.ICML 2026
- Fair regression via plug-in estimator and recalibration with statistical guaranteesEvgenii Chzhen, Christophe Denis, Mohamed Hebiri, Luca Oneto et al.NeurIPS 2020 · 52 citations
