Fixes That Fail: Self-Defeating Improvements in Machine-Learning Systems
Ruihan Wu, Chuan Guo, Awni Y. Hannun, Laurens van der Maaten
Abstract
Machine-learning systems such as self-driving cars or virtual assistants are composed of a large number of machine-learning models that recognize image content, transcribe speech, analyze natural language, infer preferences, rank options, etc. Models in these systems are often developed and trained independently, which raises an obvious concern: Can improving a machine-learning model make the overall system worse? We answer this question affirmatively by showing that improving a model can deteriorate the performance of downstream models, even after those downstream models are retrained. Such self-defeating improvements are the result of entanglement between the models in the system. We perform an error decomposition of systems with multiple machine-learning models, which sheds light on the types of errors that can lead to self-defeating improvements. We also present the results of experiments which show that self-defeating improvements emerge in a realistic stereo-based detection system for cars and pedestrians.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d720234d-37f0-449c-84c7-34d7c324151aCited by top-tier papers3
- TaskMet: Task-driven Metric Learning for Model LearningDishank Bansal, Ricky T. Q. Chen, Mustafa Mukadam, Brandon AmosNeurIPS 2023 · 20 citations
- RegTrieve: Reducing System-Level Regression Errors for Machine Learning Systems via Retrieval-Enhanced EnsembleJunming Cao, Xuwen Xiang, Mingfei Cheng, Bihuan Chen et al.FSE 2025
- Comfrey: Mitigating Integration Failures in LLM-enabled Software at Run-TimeYuchen Shao, Yuheng Huang, Jiazhen Zou, Yuling Shi et al.ICSE 2026
Builds on3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Towards Backward-Compatible Representation LearningYantao Shen, Yuanjun Xiong, Wei Xia, Stefano SoattoCVPR 2020
- Positive-Congruent Training: Towards Regression-Free Model UpdatesSijie Yan, Yuanjun Xiong, Kaustav Kundu, Shuo Yang et al.CVPR 2021
Related papers
- Towards AutoAI: Optimizing a Machine Learning System with Black-box and Differentiable ComponentsZhiliang Chen, Chuan-Sheng Foo, Bryan Kian Hsiang LowICML 2024 · 10 citations
- Joint Training of Deep Ensembles Fails Due to Learner CollusionAlan Jeffares, Tennison Liu, Jonathan Crabbé, Mihaela van der SchaarNeurIPS 2023 · 34 citations
- Ecosystem-level Analysis of Deployed Machine Learning Reveals Homogeneous OutcomesConnor Toups, Rishi Bommasani, Kathleen Creel, Sarah H. Bana et al.NeurIPS 2023 · 25 citations
- It Takes Two to EntangleZhanghan Wang, Ding Ding, Hang Zhu, Haibin Lin et al.ASPLOS 2026
- Among Us: Measuring and Mitigating Malicious Contributions in Model Collaboration SystemsZiyuan Yang, Wenxuan Ding, Shangbin Feng, Yulia TsvetkovACL 2026 · 1 citation
