← all papers · overview

Knowledge Sanitization Of Large Language Models

Abstract

We explore a knowledge sanitization approach to mitigate the privacy concerns associated with large language models (LLMs). LLMs trained on a large corpus of Web data can memorize and potentially reveal sensitive or confidential information, raising critical security concerns. Our technique efficiently fine-tunes these models using the Low-Rank Adaptation (LoRA) method, prompting them to generate

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).