Ecosystem-level Analysis of Deployed Machine Learning Reveals Homogeneous Outcomes
Connor Toups, Rishi Bommasani, Kathleen Creel, Sarah H. Bana, Dan Jurafsky, Percy Liang
Abstract
Machine learning is traditionally studied at the model level: researchers measure and improve the accuracy, robustness, bias, efficiency, and other dimensions of specific models. In practice, however, the societal impact of any machine learning model depends on the context into which it is deployed. To capture this, we introduce ecosystem-level analysis: rather than analyzing a single model, we consider the collection of models that are deployed in a given context. For example, ecosystem-level analysis in hiring recognizes that a job candidate's outcomes are determined not only by a single hiring algorithm or firm but instead by the collective decisions of all the firms to which the candidate applied. Across three modalities (text, images, speech) and eleven datasets, we establish a clear trend: deployed machine learning is prone to systemic failure, meaning some users are exclusively misclassified by all models available. Even when individual models improve over time, we find these improvements rarely reduce the prevalence of systemic failure. Instead, the benefits of these improvements predominantly accrue to individuals who are already correctly classified by other models. In light of these trends, we analyze medical imaging for dermatology, a setting where the costs of systemic failure are especially high. While traditional analyses reveal that both models and humans exhibit racial performance disparities, ecosystem-level analysis reveals new forms of racial disparity in model predictions that do not present in human predictions. These examples demonstrate that ecosystem-level analysis has unique strengths in characterizing the societal impact of machine learning. 1 * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8e653dfe-e984-48a5-a335-33bbe87adbdcCited by top-tier papers4
- Human Expertise in Algorithmic PredictionRohan Alur, Manish Raghavan, Devavrat ShahNeurIPS 2024 · 18 citations
- Monoculture in Matching MarketsKenny Peng, Nikhil GargNeurIPS 2024 · 13 citations
- Monoculture or Multiplicity: Which Is It?Mila Gorecki, Moritz HardtNeurIPS 2025 · 6 citations
- Correlated Errors in Large Language ModelsElliot Myunghoon Kim, Avi Garg, Kenny Peng, Nikhil GargICML 2025
Builds on4
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 436 citations
- Picking on the Same Person: Does Algorithmic Monoculture lead to Outcome Homogenization?Rishi Bommasani, Kathleen A. Creel, Ananya Kumar, Dan Jurafsky et al.NeurIPS 2022 · 179 citations
- How Did the Model Change? Efficiently Assessing Machine Learning API ShiftsLingjiao Chen, Matei Zaharia, James ZouICLR 2022 · 5 citations
Related papers
- Fairness under CompetitionRonen Gradwohl, Eilam Shapira, Moshe TennenholtzNeurIPS 2025 · 2 citations
- Fixes That Fail: Self-Defeating Improvements in Machine-Learning SystemsRuihan Wu, Chuan Guo, Awni Y. Hannun, Laurens van der MaatenNeurIPS 2021 · 13 citations
- The Boundaries of Fair AI in Medical Image Prognosis: A Causal PerspectiveThai-Hoang Pham, Jiayuan Chen, Seungyeon Lee, Yuanlong Wang et al.NeurIPS 2025 · 3 citations
- Improved Bayes Risk Can Yield Reduced Social Welfare Under CompetitionMeena Jagadeesan, Michael I. Jordan, Jacob Steinhardt, Nika HaghtalabNeurIPS 2023 · 20 citations
- From Plane Crashes to Algorithmic Harm: Applicability of Safety Engineering Frameworks for Responsible MLShalaleh Rismani, Renee Shelby, Andrew Smart, Edgar W. Jatho III et al.CHI 2023 · 33 citations
