The Utility and Complexity of In- and Out-of-Distribution Machine Unlearning
Youssef Allouah, Joshua Kazdan, Rachid Guerraoui, Sanmi Koyejo
Abstract
Machine unlearning, the process of selectively removing data from trained models, is increasingly crucial for addressing privacy concerns and knowledge gaps post-deployment. Despite this importance, existing approaches are often heuristic and lack formal guarantees. In this paper, we analyze the fundamental utility, time, and space complexity trade-offs of approximate unlearning, providing rigorous certification analogous to differential privacy. For in-distribution forget data -- data similar to the retain set -- we show that a surprisingly simple and general procedure, empirical risk minimization with output perturbation, achieves tight unlearning-utility-complexity trade-offs, addressing a previous theoretical gap on the separation from unlearning "for free" via differential privacy, which inherently facilitates the removal of such data. However, such techniques fail with out-of-distribution forget data -- data significantly different from the retain set -- where unlearning time complexity can exceed that of retraining, even for a single sample. To address this, we propose a new robust and noisy gradient descent variant that provably amortizes unlearning time complexity without compromising utility.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bcb3f451-a96f-4956-adcc-4253583c34e8Cited by top-tier papers9
- Gaussian certified unlearning in high dimensions: A hypothesis testing approachAaradhya Pandey, Arnab Auddy, Haolin Zou, Arian Maleki et al.ICLR 2026 · 5 citations
- Distributional Machine Unlearning via Selective Data RemovalYoussef Allouah, Rachid Guerraoui, Sanmi KoyejoICLR 2026 · 5 citations
- The Unseen Threat: Residual Knowledge in Machine Unlearning under Perturbed SamplesHsiang Hsu, Pradeep Niroula, Zichang He, Ivan Brugere et al.NeurIPS 2025 · 5 citations
- Fully Decentralized Certified UnlearningHithem Lamri, Michail ManiatakosCVPR 2026 · 1 citation
- Exact Unlearning in Reinforcement LearningTang Thanh Nguyen, Raman AroraICML 2026
Builds on12
- Machine UnlearningLucas Bourtoule, Varun Chandrasekaran, Christopher A. Choquette-Choo, Hengrui Jia et al.S&P 2021 · 1,381 citations
- Certified Data Removal from Machine Learning ModelsChuan Guo, Tom Goldstein, Awni Y. Hannun, Laurens van der MaatenICML 2020 · 633 citations
- Remember What You Want to Forget: Algorithms for Machine UnlearningAyush Sekhari, Jayadev Acharya, Gautam Kamath, Ananda Theertha SureshNeurIPS 2021 · 516 citations
- Amnesiac Machine LearningLaura Graves, Vineel Nagisetty, Vijay GaneshAAAI 2021 · 416 citations
- Towards Unbounded Machine UnlearningMeghdad Kurmanji, Peter Triantafillou, Jamie Hayes, Eleni TriantafillouNeurIPS 2023 · 363 citations
Related papers
- Rewind-to-Delete: Certified Machine Unlearning for Nonconvex FunctionsSiqiao Mu, Diego KlabjanNeurIPS 2025 · 22 citations
- Certified Unlearning for Neural NetworksAnastasia Koloskova, Youssef Allouah, Animesh Jha, Rachid Guerraoui et al.ICML 2025
- Langevin Unlearning: A New Perspective of Noisy Gradient Descent for Machine UnlearningEli Chien, Haoyu Wang, Ziang Chen, Pan LiNeurIPS 2024 · 58 citations
- Certified Machine Unlearning via Noisy Stochastic Gradient DescentEli Chien, Haoyu Wang, Ziang Chen, Pan LiNeurIPS 2024 · 16 citations
- When to Forget? Complexity Trade-offs in Machine UnlearningMartin Van Waerebeke, Marco Lorenzi, Giovanni Neglia, Kevin ScamanICML 2025
