LOTUS: Evolving Multimodal Unlearning via Hyperbolic Entailment and Lorentz Transport
Zekun Wang, Jingjie Zeng, Yingxu Li, Hongfei Lin, Liang Yang
摘要
Multimodal Large Language Models (MLLMs) face critical privacy challenges arising from the indiscriminate memorization of sensitive data. Existing unlearning methods often fail to precisely disentangle specific instances from general concepts, leading to either catastrophic forgetting of useful knowledge or unsafe content substitution. We attribute these failures to a fundamental geometric mismatch: these approaches primarily operate in Euclidean space, which lacks the capacity to model the hierarchical entailment inherent in visual-linguistic concepts. To address this, we introduce LOTUS (LOrentz Transport for Unlearning Strategies), a framework that performs surgical semantic pruning within the Lorentz manifold. LOTUS employs an Inverted Entailment Cone Loss to sever the semantic inheritance of sensitive concepts and a Lorentz Transport mechanism to align pruned features with a safety refusal prior in the tangent space. Extensive experiments on MLLMU-Bench demonstrate that LOTUS significantly outperforms baselines, improving unlearning efficacy by over 9% on LLaVA compared to state-of-the-art constraint-based methods. Crucially, LOTUS achieves this precision while maintaining general utility, effectively resolving the dilemma between thorough erasure and model stability.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- The WMDP Benchmark: Measuring and Reducing Malicious Use with UnlearningNathaniel Li, Alexander Pan, Anjali Gopal, Summer Yue 等ICML 2024 · 被引用 390 次
- Variational Bayesian UnlearningQuoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick JailletNeurIPS 2020 · 被引用 198 次
- Hyperbolic Image-text RepresentationsKaran Desai, Maximilian Nickel, Tanmay Rajpurohit, Justin Johnson 等ICML 2023 · 被引用 137 次
- Editing models with task arithmeticGabriel Ilharco, Marco Túlio Ribeiro, Mitchell Wortsman, Ludwig Schmidt 等ICLR 2023 · 被引用 31 次
- DEPN: Detecting and Editing Privacy Neurons in Pretrained Language ModelsXinwei Wu, Junzhuo Li, Minghui Xu, Weilong Dong 等EMNLP 2023 · 被引用 28 次
相关 Paper
- Modality-Aware Neuron Pruning for Unlearning in Multimodal Large Language ModelsZheyuan Liu, Guangyao Dou, Xiangchi Yuan, Chunhui Zhang 等ACL 2025
- Constrained Entropic Unlearning: A Primal-Dual Framework for Large Language ModelsTaha Entesari, Arman Hatami, Rinat Khaziev, Anil Ramakrishna 等NeurIPS 2025 · 被引用 12 次
- VL-Eraser: Vacuum Distillation for Machine Unlearning in Vision-Language ModelsYili Wang, Lu Dai, Tairan Huang, Yijie Xu 等CVPR 2026
- Towards Robust and Parameter-Efficient Knowledge Unlearning for LLMsSungmin Cha, Sungjun Cho, Dasol Hwang, Moontae LeeICLR 2025
- Unified Parameter-Efficient Unlearning for LLMsChenlu Ding, Jiancan Wu, Yancheng Yuan, Jinda Lu 等ICLR 2025
