Lune

ICML2023顶会

On Provable Copyright Protection for Generative Models

Nikhil Vyas, Sham M. Kakade, Boaz Barak

2023年份
120被引次数
34顶会引用

摘要

There is a growing concern that learned conditional generative models may output samples that are substantially similar to some copyrighted data CC that was in their training set. We give a formal definition of near access-freeness (NAF)\textit{near access-freeness (NAF)} and prove bounds on the probability that a model satisfying this definition outputs a sample similar to CC, even if CC is included in its training set. Roughly speaking, a generative model pp is \textit{k-NAF} if for every potentially copyrighted data CC, the output of pp diverges by at most kk-bits from the output of a model qq that \textit{did not access C at all}. We also give generative model learning algorithms, which efficiently modify the original generative model learning algorithm in a black box manner, that output generative models with strong bounds on the probability of sampling protected content. Furthermore, we provide promising experiments for both language (transformers) and image (diffusion) generative models, showing minimal degradation in output quality while ensuring strong protections against sampling protected content.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext d0b457fd-9b5d-48c8-8d7a-709ff1079748

引用它的顶会 Paper34

问问它们各自怎么用它

它引用的顶会 Paper11

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖