How Did the Model Change? Efficiently Assessing Machine Learning API Shifts
Lingjiao Chen, Matei Zaharia, James Zou
摘要
ML prediction APIs from providers like Amazon and Google have made it simple to use ML in applications. A challenge for users is that such APIs continuously change over time as the providers update models, and changes can happen silently without users knowing. It is thus important to monitor when and how much the ML APIs' performance shifts. To provide detailed change assessment, we model ML API shifts as confusion matrix differences, and propose a principled algorithmic framework, MASA, to provably assess these shifts efficiently given a sample budget constraint. MASA employs an upper-confidence bound based approach to adaptively determine on which data point to query the ML API to estimate shifts. Empirically, we observe significant ML API shifts from 2020 to 2021 among 12 out of 36 applications using commercial APIs from Google, Microsoft, Amazon, and other providers. These real-world shifts include both improvements and reductions in accuracy. Extensive experiments show that MASA can estimate such API shifts more accurately than standard approaches given the same budget.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Estimating and Explaining Model Performance When Both Covariates and Labels ShiftLingjiao Chen, Matei Zaharia, James Y. ZouNeurIPS 2022 · 被引用 34 次
- Ecosystem-level Analysis of Deployed Machine Learning Reveals Homogeneous OutcomesConnor Toups, Rishi Bommasani, Kathleen Creel, Sarah H. Bana 等NeurIPS 2023 · 被引用 25 次
- Efficient Online ML API Selection for Multi-Label Classification TasksLingjiao Chen, Matei Zaharia, James ZouICML 2022 · 被引用 22 次
- Log Probability Tracking of LLM APIsTimothee Chauvin, Erwan Le Merrer, Francois Taiani, Gilles TredanICLR 2026 · 被引用 12 次
- Token-Efficient Change Detection in LLM APIsTimothee Chauvin, Clément Lalanne, Erwan Le Merrer, Jean-Michel Loubes 等ICML 2026 · 被引用 4 次
它引用的顶会 Paper3
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- From ImageNet to Image Classification: Contextualizing Progress on BenchmarksDimitris Tsipras, Shibani Santurkar, Logan Engstrom, Andrew Ilyas 等ICML 2020 · 被引用 146 次
- Beyond Accuracy: Behavioral Testing of NLP Models with CheckListMarco Túlio Ribeiro, Tongshuang Wu, Carlos Guestrin, Sameer SinghACL 2020 · 被引用 51 次
相关 Paper
- FrugalML: How to use ML Prediction APIs more accurately and cheaplyLingjiao Chen, Matei Zaharia, James Y. ZouNeurIPS 2020 · 被引用 57 次
- Are Machine Learning Cloud APIs Used Correctly?Chengcheng Wan, Shicheng Liu, Henry Hoffmann, Michael Maire 等ICSE 2021 · 被引用 37 次
- Model Equality Testing: Which Model is this API Serving?Irena Gao, Percy Liang, Carlos GuestrinICLR 2025
- ChameleonAPI: Automatic and Efficient Customization of Neural Networks for ML ApplicationsYuhan Liu, Chengcheng Wan, Kuntai Du, Henry Hoffmann 等OSDI 2024 · 被引用 1 次
- Run-Time Prevention of Software Integration Failures of Machine Learning APIsChengcheng Wan, Yuhan Liu, Kuntai Du, Henry Hoffmann 等OOPSLA 2023 · 被引用 5 次
