← all papers · overview

Improving Sparse Memory Finetuning

Abstract

Large Language Models (LLMs) are typically static after training, yet real-world applications require continual adaptation to new knowledge without degrading existing capabilities. Standard approaches to updating models, like full finetuning or parameter-efficient methods (e.g., LoRA), face a fundamental trade-off: catastrophic forgetting. They modify shared dense representations, causing interfer

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).