Backward-Compatible Prediction Updates: A Probabilistic Approach
Frederik Träuble, Julius von Kügelgen, Matthäus Kleindessner, Francesco Locatello, Bernhard Schölkopf, Peter V. Gehler
摘要
When machine learning systems meet real world applications, accuracy is only one of several requirements. In this paper, we assay a complementary perspective originating from the increasing availability of pre-trained and regularly improving state-of-the-art models. While new improved models develop at a fast pace, downstream tasks vary more slowly or stay constant. Assume that we have a large unlabelled data set for which we want to maintain accurate predictions. Whenever a new and presumably better ML models becomes available, we encounter two problems: (i) given a limited budget, which data points should be re-evaluated using the new model?; and (ii) if the new predictions differ from the current ones, should we update? Problem (i) is about compute cost, which matters for very large data sets and models. Problem (ii) is about maintaining consistency of the predictions, which can be highly relevant for downstream applications; our demand is to avoid negative flips, i.e., changing correct to incorrect predictions. In this paper, we formalize the Prediction Update Problem and present an efficient probabilistic approach as answer to the above questions. In extensive experiments on standard classification benchmark data sets, we show that our method outperforms alternative strategies along key metrics for backward-compatible prediction updates.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Measuring and Reducing Model Update Regression in Structured Prediction for NLPDeng Cai, Elman Mansimov, Yi-An Lai, Yixuan Su 等NeurIPS 2022 · 被引用 14 次
- BT2: Backward-compatible Training with Basis TransformationYifei Zhou, Zilu Li, Abhinav Shrivastava, Hengshuang Zhao 等ICCV 2023 · 被引用 7 次
- Lightweight Approaches to DNN Regression Error Reduction: An Uncertainty Alignment PerspectiveZenan Li, Maorun Zhang, Jingwei Xu, Yuan Yao 等ICSE 2023 · 被引用 3 次
- RegTrieve: Reducing System-Level Regression Errors for Machine Learning Systems via Retrieval-Enhanced EnsembleJunming Cao, Xuwen Xiang, Mingfei Cheng, Bihuan Chen 等FSE 2025
- Mitigating Negative Flips via Margin Preserving TrainingSimone Ricci, Niccolò Biondi, Federico Pernici, Alberto Del BimboAAAI 2026
它引用的顶会 Paper5
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 被引用 1,861 次
- Towards Backward-Compatible Representation LearningYantao Shen, Yuanjun Xiong, Wei Xia, Stefano SoattoCVPR 2020
- Positive-Congruent Training: Towards Regression-Free Model UpdatesSijie Yan, Yuanjun Xiong, Kaustav Kundu, Shuo Yang 等CVPR 2021
- Self-Training With Noisy Student Improves ImageNet ClassificationQizhe Xie, Minh-Thang Luong, Eduard H. Hovy, Quoc V. LeCVPR 2020
相关 Paper
- When to retrain a machine learning modelFlorence Regol, Leo Schwinn, Kyle Sprague, Mark Coates 等ICML 2025
- A Generalized Backward Compatibility MetricTomoya SakaiKDD 2022 · 被引用 3 次
- Hot-Refresh Model Upgrades with Regression-Free Compatible Training in Image RetrievalBinjie Zhang, Yixiao Ge, Yantao Shen, Yu Li 等ICLR 2022 · 被引用 13 次
- Contextual Active Model SelectionXuefeng Liu, Fangfang Xia, Rick Stevens, Yuxin ChenNeurIPS 2024
- FastFill: Efficient Compatible Model UpdateFlorian Jaeckle, Fartash Faghri, Ali Farhadi, Oncel Tuzel 等ICLR 2023
