Fairness Evaluation in Text Classification: Machine Learning Practitioner Perspectives of Individual and Group Fairness
Zahra Ashktorab, Benjamin Hoover, Mayank Agarwal, Casey Dugan, Werner Geyer, Hao Bang Yang, Mikhail Yurochkin
摘要
Mitigating algorithmic bias is a critical task in the development and deployment of machine learning models. While several toolkits exist to aid machine learning practitioners in addressing fairness issues, little is known about the strategies practitioners employ to evaluate model fairness and what factors influence their assessment, particularly in the context of text classification. Two common approaches of evaluating the fairness of a model are group fairness and individual fairness. We run a study with Machine Learning practitioners (n=24) to understand the strategies used to evaluate models. Metrics presented to practitioners (group vs. individual fairness) impact which models they consider fair. Participants focused on risks associated with underpredicting / overpredicting and model sensitivity relative to identity token manipulations. We discover fairness assessment strategies involving personal experiences or how users form groups of identity tokens to test model fairness. We provide recommendations for interactive tools for evaluating fairness in text classification.
• Human-centered computing → Human computer interaction (HCI).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Towards a Non-Ideal Methodological Framework for Responsible MLRamaravind Kommiya Mothilal, Shion Guha, Syed Ishtiaque AhmedCHI 2024 · 被引用 7 次
- Data Ethics Emergency Drill: A Toolbox for Discussing Responsible AI for Industry TeamsVanessa Aisyahsari Hanschke, Dylan Rees, Merve Alanyali, David Hopkinson 等CHI 2024 · 被引用 7 次
- EARN Fairness: Explaining, Asking, Reviewing, and Negotiating Artificial Intelligence Fairness Metrics Among StakeholdersLin Luo, Yuri Nakao, Mathieu Chollet, Hiroya Inakoshi 等CSCW 2025 · 被引用 4 次
- "I think this is fair": Uncovering the Complexities of Stakeholder Decision-Making in AI Fairness AssessmentLin Luo, Yuri Nakao, Mathieu Chollet, Hiroya Inakoshi 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper11
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Co-Designing Checklists to Understand Organizational Challenges and Opportunities around Fairness in AIMichael A. Madaio, Luke Stark, Jennifer Wortman Vaughan, Hanna M. WallachCHI 2020 · 被引用 428 次
- Assessing the Fairness of AI Systems: AI Practitioners' Processes, Challenges, and Needs for SupportMichael Madaio, Lisa Egede, Hariharan Subramonyam, Jennifer Wortman Vaughan 等CSCW 2022 · 被引用 149 次
- Training individually fair ML models with sensitive subspace robustnessMikhail Yurochkin, Amanda Bower, Yuekai SunICLR 2020 · 被引用 123 次
- Post-processing for Individual FairnessFelix Petersen, Debarghya Mukherjee, Yuekai Sun, Mikhail YurochkinNeurIPS 2021 · 被引用 115 次
相关 Paper
- Towards Fairness in Practice: A Practitioner-Oriented Rubric for Evaluating Fair ML ToolkitsBrianna Richardson, Jean Garcia-Gathright, Samuel F. Way, Jennifer Thom 等CHI 2021 · 被引用 56 次
- The Landscape and Gaps in Open Source Fairness ToolkitsMichelle Seng Ah Lee, Jatinder SinghCHI 2021 · 被引用 117 次
- Do the machine learning models on a crowd sourced platform exhibit bias? an empirical study on model fairnessSumon Biswas, Hridesh RajanFSE 2020 · 被引用 96 次
- Two Simple Ways to Learn Individual Fairness Metrics from DataDebarghya Mukherjee, Mikhail Yurochkin, Moulinath Banerjee, Yuekai SunICML 2020 · 被引用 109 次
- OxonFair: A Flexible Toolkit for Algorithmic FairnessEoin Delaney, Zihao Fu, Sandra Wachter, Brent D. Mittelstadt 等NeurIPS 2024 · 被引用 13 次
