← all papers · overview

Measuring Copyright Risks Of Large Language Model Via Partial Information Probing

Abstract

Exploring the data sources used to train Large Language Models (LLMs) is a crucial direction in investigating potential copyright infringement by these models. While this approach can identify the possible use of copyrighted materials in training data, it does not directly measure infringing risks. Recent research has shifted towards testing whether LLMs can directly output copyrighted content. Ad

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).