← all papers · overview

RWKU: Benchmarking Real-world Knowledge Unlearning For Large Language Models

Abstract

Large language models (LLMs) inevitably memorize sensitive, copyrighted, and harmful knowledge from the training corpus; therefore, it is crucial to erase this knowledge from the models. Machine unlearning is a promising solution for efficiently removing specific knowledge by post hoc modifying models. In this paper, we propose a Real-World Knowledge Unlearning benchmark (RWKU) for LLM unlearning.

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).