← all papers · overview

Quantization-robust LLM Unlearning Via Low-rank Adaptation

Abstract

Large Language Model (LLM) unlearning aims to remove targeted knowledge from a trained model, but practical deployments often require post-training quantization (PTQ) for efficient inference. However, aggressive low-bit PTQ can mask unlearning updates, causing quantized models to revert to pre-unlearning behavior. We show that standard full-parameter fine-tuning often induces parameter changes tha

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).