Exact Unlearning in Reinforcement Learning
Tang Thanh Nguyen, Raman Arora
摘要
We formulate the problem of exact unlearning in reinforcement learning, where the goal is to design an efficient framework that enables the removal of any user’s data upon deletion request, i.e., the online learner’s output after unlearning be indistinguishable from what would have been produced had the deleted user never interacted with the learner. For any , we show that there exists a reinforcement learning (RL) algorithm that is -TV-stable and supports an exact unlearning procedure whose expected computational cost is only a fraction of the computational cost of retraining from scratch. We construct such a -TV-stable RL algorithm for tabular Markov decision processes (MDPs), which achieves a regret bound of , where , and denote the number of states, the number of actions, the episode horizon, and the number of episodes, respectively. We also establish a lower bound of for -TV-stable RL algorithms, showing that our algorithm is nearly minimax optimal.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Machine UnlearningLucas Bourtoule, Varun Chandrasekaran, Christopher A. Choquette-Choo, Hengrui Jia 等S&P 2021 · 被引用 1,381 次
- Certified Data Removal from Machine Learning ModelsChuan Guo, Tom Goldstein, Awni Y. Hannun, Laurens van der MaatenICML 2020 · 被引用 633 次
- Remember What You Want to Forget: Algorithms for Machine UnlearningAyush Sekhari, Jayadev Acharya, Gautam Kamath, Ananda Theertha SureshNeurIPS 2021 · 被引用 516 次
- Differentially Private Regret Minimization in Episodic Markov Decision ProcessesSayak Ray Chowdhury, Xingyu ZhouAAAI 2022 · 被引用 26 次
- The Utility and Complexity of In- and Out-of-Distribution Machine UnlearningYoussef Allouah, Joshua Kazdan, Rachid Guerraoui, Sanmi KoyejoICLR 2025
相关 Paper
- Communication Efficient and Provable Federated UnlearningYouming Tao, Cheng-Long Wang, Miao Pan, Dongxiao Yu 等VLDB 2024 · 被引用 35 次
- Algorithms that Approximate Data Removal: New Results and LimitationsVinith M. Suriyakumar, Ashia C. WilsonNeurIPS 2022 · 被引用 55 次
- Hard to Forget: Poisoning Attacks on Certified Machine UnlearningNeil G. Marchant, Benjamin I. P. Rubinstein, Scott AlfeldAAAI 2022 · 被引用 95 次
- Hessian-Free Online Certified UnlearningXinbao Qiao, Meng Zhang, Ming Tang, Ermin WeiICLR 2025
- When to Forget? Complexity Trade-offs in Machine UnlearningMartin Van Waerebeke, Marco Lorenzi, Giovanni Neglia, Kevin ScamanICML 2025
