← all papers · overview

Dissecting Language Models: Machine Unlearning Via Selective Pruning

Abstract

Understanding and shaping the behaviour of Large Language Models (LLMs) is increasingly important as applications become more powerful and more frequently adopted. This paper introduces a machine unlearning method specifically designed for LLMs. We introduce a selective pruning method for LLMs that removes neurons based on their relative importance on a targeted capability compared to overall netw

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).