← all papers · overview

Measuring Memorization In Language Models Via Probabilistic Extraction

Abstract

Large language models (LLMs) are susceptible to memorizing training data, raising concerns about the potential extraction of sensitive information at generation time. Discoverable extraction is the most common method for measuring this issue: split a training example into a prefix and suffix, then prompt the LLM with the prefix, and deem the example extractable if the LLM generates the matching su

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).