← all papers · overview

Scaling Sparse Fine-tuning To Large Language Models

Abstract

Large Language Models (LLMs) are difficult to fully fine-tune (e.g., with instructions or human feedback) due to their sheer number of parameters. A family of parameter-efficient sparse fine-tuning methods have proven promising in terms of performance but their memory requirements increase proportionally to the size of the LLMs. In this work, we scale sparse fine-tuning to state-of-the-art LLMs li

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).