← all papers · overview

Rethinking LLM Unlearning Objectives: A Gradient Perspective And Go Beyond

Abstract

Large language models (LLMs) should undergo rigorous audits to identify potential risks, such as copyright and privacy infringements. Once these risks emerge, timely updates are crucial to remove undesirable responses, ensuring legal and safe model usage. It has spurred recent research into LLM unlearning, focusing on erasing targeted undesirable knowledge without compromising the integrity of oth

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).