SAFE: Machine Unlearning With Shard Graphs
Yonatan Dukler, Benjamin Bowman, Alessandro Achille, Aditya Golatkar, Ashwin Swaminathan, Stefano Soatto
摘要
We present Synergy Aware Forgetting Ensemble (SAFE), a method to adapt large models on a diverse collection of data while minimizing the expected cost to remove the influence of training samples from the trained model. This process, also known as selective forgetting or unlearning, is often conducted by partitioning a dataset into shards, training fully independent models on each, then ensembling the resulting models. Increasing the number of shards reduces the expected cost to forget but at the same time it increases inference cost and reduces the final accuracy of the model since synergistic information between samples is lost during the independent model training. Rather than treating each shard as independent, SAFE introduces the notion of a shard graph, which allows incorporating limited information from other shards during training, trading off a modest increase in expected forgetting cost with a significant increase in accuracy, all while still attaining complete removal of residual influence after forgetting. SAFE uses a lightweight system of adapters which can be trained while reusing most of the computations. This allows SAFE to be trained on shards an order-of-magnitude smaller than current state-of-the-art methods (thus reducing the forgetting costs) while also maintaining high accuracy, as we demonstrate empirically on fine-grained computer vision datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- On Effects of Steering Latent Representation for Large Language Model UnlearningHuu-Tien Dang, Tin Pham, Hoang Thanh-Tung, Naoya InoueAAAI 2025 · 被引用 33 次
- Sample Selection via Contrastive Fragmentation for Noisy Label RegressionChris Dongjoo Kim, Sangwoo Moon, Jihwan Moon, Dongyeon Woo 等NeurIPS 2024 · 被引用 8 次
- Pre-training for Recommendation UnlearningGuoxuan Chen, Lianghao Xia, Chao HuangSIGIR 2025 · 被引用 2 次
- Towards Scalable Exact Machine Unlearning Using Parameter-Efficient Fine-TuningSomnath Basu Roy Chowdhury, Krzysztof Marcin Choromanski, Arijit Sehanobish, Kumar Avinava Dubey 等ICLR 2025 · 被引用 1 次
- FedShard: Federated Unlearning with Efficiency Fairness and Performance FairnessSiyuan Wen, Meng Zhang, Yang Yang, Ningning DingAAAI 2026 · 被引用 1 次
它引用的顶会 Paper19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan 等CCS 2016 · 被引用 7,620 次
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang 等ICCV 2019 · 被引用 2,239 次
- Machine UnlearningLucas Bourtoule, Varun Chandrasekaran, Christopher A. Choquette-Choo, Hengrui Jia 等S&P 2021 · 被引用 1,381 次
- Semi-Supervised Domain Adaptation via Minimax EntropyKuniaki Saito, Donghyun Kim, Stan Sclaroff, Trevor Darrell 等ICCV 2019 · 被引用 725 次
相关 Paper
- AUTE: Peer-Alignment and Self-Unlearning Boost Adversarial Robustness for Training Ensemble ModelsLifeng Huang, Tian Su, Chengying Gao, Ning Liu 等AAAI 2025 · 被引用 2 次
- PAGE: A Unified Approach for Federated Graph UnlearningYuming Ai, Xunkai Li, Jiaqi Chao, Bowen Fan 等AAAI 2026
- On the Misalignment Between Data Learnability and Forgettability in Machine UnlearningZijie Pan, Zuobin Ying, Yajie Wang, Wanlei ZhouAAAI 2026
- Targeted Forgetting of Image Subgroups in CLIP ModelsZeliang Zhang, Gaowen Liu, Charles Fleming, Ramana Rao Kompella 等CVPR 2025
- Forget What Has Seen: Selective Concept Unlearning in Segmentation Foundation ModelsMiaozeng Du, Jiaqi Li, Sirui Pan, Yi Zhan 等AAAI 2026
