← all papers · overview

Learnable Privacy Neurons Localization In Language Models

Abstract

Concerns regarding Large Language Models (LLMs) to memorize and disclose private information, particularly Personally Identifiable Information (PII), become prominent within the community. Many efforts have been made to mitigate the privacy risks. However, the mechanism through which LLMs memorize PII remains poorly understood. To bridge this gap, we introduce a pioneering method for pinpointing P

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).