← all papers · overview

Can Watermarking Large Language Models Prevent Copyrighted Text Generation And Hide Training Data?

Abstract

Large Language Models (LLMs) have demonstrated impressive capabilities in generating diverse and contextually rich text. However, concerns regarding copyright infringement arise as LLMs may inadvertently produce copyrighted material. In this paper, we first investigate the effectiveness of watermarking LLMs as a deterrent against the generation of copyrighted texts. Through theoretical analysis an

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).