← all papers · overview

Under The Surface: Tracking The Artifactuality Of Llm-generated Data

Abstract

This work delves into the expanding role of large language models (LLMs) in generating artificial data. LLMs are increasingly employed to create a variety of outputs, including annotations, preferences, instruction prompts, simulated dialogues, and free text. As these forms of LLM-generated data often intersect in their application, they exert mutual influence on each other and raise significant c

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).