Repairing Neural Networks by Leaving the Right Past Behind
Ryutaro Tanno, Melanie F. Pradier, Aditya V. Nori, Yingzhen Li
摘要
Prediction failures of machine learning models often arise from deficiencies in training data, such as incorrect labels, outliers, and selection biases. However, such data points that are responsible for a given failure mode are generally not known a priori, let alone a mechanism for repairing the failure. This work draws on the Bayesian view of continual learning, and develops a generic framework for both, identifying training examples which have given rise to the target failure, and fixing the model through erasing information about them. This framework naturally allows leveraging recent advances in continual learning to this new problem of model repairment, while subsuming the existing works on influence functions and data deletion as specific instances. Experimentally, the proposed approach outperforms the baselines for both identification of detrimental training data and fixing model failures in a generalisable manner.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Ablating Concepts in Text-to-Image Diffusion ModelsNupur Kumari, Bingliang Zhang, Sheng-Yu Wang, Eli Shechtman 等ICCV 2023 · 被引用 327 次
- Can Sensitive Information Be Deleted From LLMs? Objectives for Defending Against Extraction AttacksVaidehi Patil, Peter Hase, Mohit BansalICLR 2024 · 被引用 167 次
- Data Attribution for Text-to-Image Models by Unlearning Synthesized ImagesSheng-Yu Wang, Aaron Hertzmann, Alexei A. Efros, Jun-Yan Zhu 等NeurIPS 2024 · 被引用 28 次
- The Memory-Perturbation Equation: Understanding Model's Sensitivity to DataPeter Nickl, Lu Xu, Dharmesh Tailor, Thomas Möllenhoff 等NeurIPS 2023 · 被引用 17 次
- Redirection for Erasing Memory (REM): Towards a universal unlearning method for corrupted dataStefan Schoepf, Michael Mozer, Nicole Mitchell, Alexandra Brintrup 等ICLR 2026 · 被引用 7 次
它引用的顶会 Paper10
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Certified Data Removal from Machine Learning ModelsChuan Guo, Tom Goldstein, Awni Y. Hannun, Laurens van der MaatenICML 2020 · 被引用 633 次
- Adaptive Machine UnlearningVarun Gupta, Christopher Jung, Seth Neel, Aaron Roth 等NeurIPS 2021 · 被引用 262 次
- Editable Neural NetworksAnton Sinitsin, Vsevolod Plokhotnyuk, Dmitry V. Pyrkin, Sergei Popov 等ICLR 2020 · 被引用 210 次
- Variational Bayesian UnlearningQuoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick JailletNeurIPS 2020 · 被引用 198 次
相关 Paper
- Retaining Beneficial Information from Detrimental Data for Neural Network RepairLong-Kai Huang, Peilin Zhao, Junzhou Huang, Sinno Jialin PanNeurIPS 2023 · 被引用 2 次
- Data Glitches Discovery using Influence-based Model ExplanationsNikolaos Myrtakis, Ioannis Tsamardinos, Vassilis ChristophidesKDD 2025
- ErrorEraser: Unlearning Data Bias for Improved Continual LearningXuemei Cao, Hanlin Gu, Xin Yang, Bingjun Wei 等KDD 2025 · 被引用 1 次
- DeMix: Debugging Training Data with Mixed Data Error Types by Investigating Influence VectorsJiale Deng, Yanyan Shen, Xiaogang Shi, Junjun ChaiKDD 2026
- Machine Unlearning of Features and LabelsAlexander Warnecke, Lukas Pirch, Christian Wressnegger, Konrad RieckNDSS 2023
