Lune

ICML2026顶会

Copyright-Bench: Agentic Evaluation of Copyright Law Compliance

Zheng Hui, Doni Bloomfield, Noam Kolt

2026年份
1被引次数

摘要

Large language model (LLM) agents increasingly perform commercial tasks that involve retrieving external content such as images and, where appropriate, reproducing that content. LLM agents should comply with the law, including copyright law. Presently, however, we lack adequate frameworks to assess whether they do so in practice. To that end, we introduce Copyright-Bench, a benchmark designed to evaluate LLM agents' compliance with copyright law. Copyright-Bench is comprised of realistic commercial taskswebsite development, merchandise design, and pitch deck production-that involve agents selecting between public-domain content (the use of which is legal) and copyrighted content (the use of which is infringing in this setting). The evaluation introduces prompt variations that simulate different user preferences, as well as time pressure. Comparing state-of-the-art LLM agents against a human baseline, we find that: (1) agents select copyrighted works despite the availability of public-domain alternatives; and (2) for openweights models, violation rates increase in response to certain user preferences and simulated time pressure.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper12

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖