Ecosystem-level Analysis of Deployed Machine Learning Reveals Homogeneous Outcomes
Connor Toups, Rishi Bommasani, Kathleen Creel, Sarah H. Bana, Dan Jurafsky, Percy Liang
摘要
Machine learning is traditionally studied at the model level: researchers measure and improve the accuracy, robustness, bias, efficiency, and other dimensions of specific models. In practice, however, the societal impact of any machine learning model depends on the context into which it is deployed. To capture this, we introduce ecosystem-level analysis: rather than analyzing a single model, we consider the collection of models that are deployed in a given context. For example, ecosystem-level analysis in hiring recognizes that a job candidate's outcomes are determined not only by a single hiring algorithm or firm but instead by the collective decisions of all the firms to which the candidate applied. Across three modalities (text, images, speech) and eleven datasets, we establish a clear trend: deployed machine learning is prone to systemic failure, meaning some users are exclusively misclassified by all models available. Even when individual models improve over time, we find these improvements rarely reduce the prevalence of systemic failure. Instead, the benefits of these improvements predominantly accrue to individuals who are already correctly classified by other models. In light of these trends, we analyze medical imaging for dermatology, a setting where the costs of systemic failure are especially high. While traditional analyses reveal that both models and humans exhibit racial performance disparities, ecosystem-level analysis reveals new forms of racial disparity in model predictions that do not present in human predictions. These examples demonstrate that ecosystem-level analysis has unique strengths in characterizing the societal impact of machine learning. 1 * Equal contribution.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Human Expertise in Algorithmic PredictionRohan Alur, Manish Raghavan, Devavrat ShahNeurIPS 2024 · 被引用 18 次
- Monoculture in Matching MarketsKenny Peng, Nikhil GargNeurIPS 2024 · 被引用 13 次
- Monoculture or Multiplicity: Which Is It?Mila Gorecki, Moritz HardtNeurIPS 2025 · 被引用 6 次
- Correlated Errors in Large Language ModelsElliot Myunghoon Kim, Avi Garg, Kenny Peng, Nikhil GargICML 2025
它引用的顶会 Paper4
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 被引用 436 次
- Picking on the Same Person: Does Algorithmic Monoculture lead to Outcome Homogenization?Rishi Bommasani, Kathleen A. Creel, Ananya Kumar, Dan Jurafsky 等NeurIPS 2022 · 被引用 179 次
- How Did the Model Change? Efficiently Assessing Machine Learning API ShiftsLingjiao Chen, Matei Zaharia, James ZouICLR 2022 · 被引用 5 次
相关 Paper
- Fairness under CompetitionRonen Gradwohl, Eilam Shapira, Moshe TennenholtzNeurIPS 2025 · 被引用 2 次
- Fixes That Fail: Self-Defeating Improvements in Machine-Learning SystemsRuihan Wu, Chuan Guo, Awni Y. Hannun, Laurens van der MaatenNeurIPS 2021 · 被引用 13 次
- The Boundaries of Fair AI in Medical Image Prognosis: A Causal PerspectiveThai-Hoang Pham, Jiayuan Chen, Seungyeon Lee, Yuanlong Wang 等NeurIPS 2025 · 被引用 3 次
- Improved Bayes Risk Can Yield Reduced Social Welfare Under CompetitionMeena Jagadeesan, Michael I. Jordan, Jacob Steinhardt, Nika HaghtalabNeurIPS 2023 · 被引用 20 次
- From Plane Crashes to Algorithmic Harm: Applicability of Safety Engineering Frameworks for Responsible MLShalaleh Rismani, Renee Shelby, Andrew Smart, Edgar W. Jatho III 等CHI 2023 · 被引用 33 次
