Assessing the Fairness of AI Systems: AI Practitioners' Processes, Challenges, and Needs for Support
Michael Madaio, Lisa Egede, Hariharan Subramonyam, Jennifer Wortman Vaughan, Hanna M. Wallach
Abstract
Various tools and practices have been developed to support practitioners in identifying, assessing, and mitigating fairness-related harms caused by AI systems. However, prior research has highlighted gaps between the intended design of these tools and practices and their use within particular contexts, including gaps caused by the role that organizational factors play in shaping fairness work. In this paper, we investigate these gaps for one such practice: disaggregated evaluations of AI systems, intended to uncover performance disparities between demographic groups. By conducting semi-structured interviews and structured workshops with thirty-three AI practitioners from ten teams at three technology companies, we identify practitioners' processes, challenges, and needs for support when designing disaggregated evaluations. We find that practitioners face challenges when choosing performance metrics, identifying the most relevant direct stakeholders and demographic groups on which to focus, and collecting datasets with which to conduct disaggregated evaluations. More generally, we identify impacts on fairness work stemming from a lack of engagement with direct stakeholders or domain experts, business imperatives that prioritize customers over marginalized groups, and the drive to deploy AI systems at scale.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 316fbb2f-46f4-4c58-807b-977f51e435b4Cited by top-tier papers45
- Seeing Like a Toolkit: How Toolkits Envision the Work of AI EthicsRichmond Y. Wong, Michael A. Madaio, Nick MerrillCSCW 2023 · 108 citations
- Investigating How Practitioners Use Human-AI Guidelines: A Case Study on the People + AI GuidebookNur Yildirim, Mahima Pushkarna, Nitesh Goyal, Martin Wattenberg et al.CHI 2023 · 103 citations
- Designing Responsible AI: Adaptations of UX Practice to Meet Responsible AI ChallengesQiaosi Wang, Michael Madaio, Shaun K. Kane, Shivani Kapania et al.CHI 2023 · 90 citations
- Designerly Understanding: Information Needs for Model Transparency to Support Design Ideation for AI-Powered User ExperienceQ. Vera Liao, Hariharan Subramonyam, Jennifer Wang, Jennifer Wortman VaughanCHI 2023 · 81 citations
- End-User Audits: A System Empowering Communities to Lead Large-Scale Investigations of Harmful Algorithmic BehaviorMichelle S. Lam, Mitchell L. Gordon, Danaë Metaxa, Jeffrey T. Hancock et al.CSCW 2022 · 77 citations
Builds on8
- Co-Designing Checklists to Understand Organizational Challenges and Opportunities around Fairness in AIMichael A. Madaio, Luke Stark, Jennifer Wortman Vaughan, Hanna M. WallachCHI 2020 · 428 citations
- Where Responsible AI meets Reality: Practitioner Perspectives on Enablers for Shifting Organizational PracticesBogdana Rakova, Jingying Yang, Henriette Cramer, Rumman ChowdhuryCSCW 2021 · 326 citations
- Between Subjectivity and Imposition: Power Dynamics in Data Annotation for Computer VisionMilagros Miceli, Martin Schuessler, Tianling YangCSCW 2020 · 148 citations
- How AI Developers Overcome Communication Challenges in a Multidisciplinary Team: A Case StudyDavid Piorkowski, Soya Park, April Yi Wang, Dakuo Wang et al.CSCW 2021 · 142 citations
- Beyond Expertise and Roles: A Framework to Characterize the Stakeholders of Interpretable Machine Learning and their NeedsHarini Suresh, Steven R. Gomez, Kevin K. Nam, Arvind SatyanarayanCHI 2021 · 115 citations
Related papers
- Understanding challenges to the interpretation of disaggregated evaluations of algorithmic fairnessStephen Pfohl, Natalie Harris, Chirag Nagpal, David Madras et al.NeurIPS 2025 · 9 citations
- Tinker, Tailor, Configure, Customize: The Articulation Work of Contextualizing an AI Fairness ChecklistMichael A. Madaio, Jingya Chen, Hanna M. Wallach, Jennifer Wortman VaughanCSCW 2024 · 13 citations
- SureMap: Simultaneous mean estimation for single-task and multi-task disaggregated evaluationMisha Khodak, Lester Mackey, Alexandra Chouldechova, Miro DudíkNeurIPS 2024 · 1 citation
- The Landscape and Gaps in Open Source Fairness ToolkitsMichelle Seng Ah Lee, Jatinder SinghCHI 2021 · 117 citations
- Zeno: An Interactive Framework for Behavioral Evaluation of Machine LearningÁngel Alexander Cabrera, Erica Fu, Donald Bertucci, Kenneth Holstein et al.CHI 2023 · 51 citations
