← all papers · overview

User Inference Attacks On Large Language Models

Abstract

Fine-tuning is a common and effective method for tailoring large language models (LLMs) to specialized tasks and applications. In this paper, we study the privacy implications of fine-tuning LLMs on user data. To this end, we consider a realistic threat model, called user inference, wherein an attacker infers whether or not a user's data was used for fine-tuning. We design attacks for performing u

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).