Lune

KDD2026顶会

LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing

Jiahao Chen, Junhao Li, Yiming Wang, Yong Yang, Yi Jiang, Chunyi Zhou, Qingming Li, Tianyu Du, Shouling Ji

2026年份
2被引次数
1顶会引用

摘要

The proliferation of Low-Rank Adaptation (LoRA) has democratized personalized text-to-image generation, enabling users to share lightweight models (e.g., personal portraits) on platforms like Civitai and Liblib. However, this ''share-and-play'' ecosystem introduces critical but unnoticed risks: benign LoRAs can be weaponized by adversaries to generate harmful content (e.g., political, defamatory imagery), undermining creator rights and platform safety. To bridge this gap, we propose LoRAShield, the first data-free editing framework for securing LoRA models against misuse. Our platformdriven approach dynamically edits and realigns LoRA's weight subspace via adversarial optimization and semantic augmentation. Experimental results demonstrate that LoRAShield achieves remarkable effectiveness, efficiency, and robustness in blocking malicious generations without sacrificing the functionality of the benign task. By shifting the defense to platforms, LoRAShield enables secure, scalable sharing of personalized models, a critical step toward trustworthy generative ecosystems. Warnings: This paper contains sexually and bloody explicit imagery that some readers may find disturbing, distressing, and/or offensive. To mitigate the offensiveness to readers, we showcase some of the explicit images with black masks.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper1

问问它们各自怎么用它

它引用的顶会 Paper22

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖