Adversarial Mixup Unlearning
Zhuoyi Peng, Yixuan Tang, Yi Yang
摘要
Machine unlearning is a critical area of research aimed at safeguarding data privacy by enabling the removal of sensitive information from machine learning models. One unique challenge in this field is catastrophic unlearning, where erasing specific data from a well-trained model unintentionally removes essential knowledge, causing the model to deviate significantly from a retrained one. To address this, we introduce a novel approach that regularizes the unlearning process by utilizing synthesized mixup samples, which simulate the data susceptible to catastrophic effects. At the core of our approach is a generator-unlearner framework, MixUnlearn, where a generator adversarially produces challenging mixup examples, and the unlearner effectively forgets target information based on these synthesized data. Specifically, we first introduce a novel contrastive objective to train the generator in an adversarial direction: generating examples that prompt the unlearner to reveal information that should be forgotten, while losing essential knowledge. Then the unlearner, guided by two other contrastive loss terms, processes the synthesized and real data jointly to ensure accurate unlearning without losing critical knowledge, overcoming catastrophic effects. Extensive evaluations across benchmark datasets demonstrate that our method significantly outperforms state-of-the-art approaches, offering a robust solution to machine unlearning. This work not only deepens understanding of unlearning mechanisms but also lays the foundation for effective machine unlearning with mixup augmentation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language ModelsHaokun Chen, Jianing Li, Yao Zhang, Jinhe Bi 等AAAI 2026
- PeriUn: Enhancing Unlearning by Selectively Forgetting Peripheral SamplesHee Bin Yoo, Dong-Sig Han, Jaein Kim, Byoung-Tak ZhangAAAI 2026
它引用的顶会 Paper9
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Membership Inference Attacks Against Machine Learning ModelsReza Shokri, Marco Stronati, Congzheng Song, Vitaly ShmatikovS&P 2017 · 被引用 5,137 次
- Towards Unbounded Machine UnlearningMeghdad Kurmanji, Peter Triantafillou, Jamie Hayes, Eleni TriantafillouNeurIPS 2023 · 被引用 363 次
- Unlearnable Examples: Making Personal Data UnexploitableHanxun Huang, Xingjun Ma, Sarah Monazam Erfani, James Bailey 等ICLR 2021 · 被引用 255 次
- Can Bad Teaching Induce Forgetting? Unlearning in Deep Networks Using an Incompetent TeacherVikram S. Chundawat, Ayush K. Tarun, Murari Mandal, Mohan S. KankanhalliAAAI 2023 · 被引用 247 次
相关 Paper
- CoUn: Empowering Machine Unlearning via Contrastive LearningYasser H. Khalil, Mehdi Setayesh, Hongliang LiNeurIPS 2025 · 被引用 4 次
- RUAGO: Effective and Practical Retain-Free Unlearning via Adversarial Attack and OOD GeneratorSangyong Lee, Sangjun Chung, Simon S. WooNeurIPS 2025 · 被引用 3 次
- Label-Agnostic Forgetting: A Supervision-Free Unlearning in Deep ModelsShaofei Shen, Chenhao Zhang, Yawen Zhao, Alina Bialkowski 等ICLR 2024 · 被引用 20 次
- Unlearning-Aware MinimizationHoki Kim, Keonwoo Kim, Sungwon Chae, Sangwon YoonNeurIPS 2025 · 被引用 7 次
- Efficient Source-free Unlearning via Energy-Guided Data Synthesis and Discrimination-Aware Multitask OptimizationXiuyuan Wang, Chaochao Chen, Weiming Liu, Xinting Liao 等ICML 2025
