← all papers · overview

Source Attribution For Large Language Model-generated Data

Abstract

The impressive performances of Large Language Models (LLMs) and their immense potential for commercialization have given rise to serious concerns over the Intellectual Property (IP) of their training data. In particular, the synthetic texts generated by LLMs may infringe the IP of the data being used to train the LLMs. To this end, it is imperative to be able to perform source attribution by ident

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).