← all papers · overview

Quantized Side Tuning: Fast And Memory-efficient Tuning Of Quantized Large Language Models

Abstract

Finetuning large language models (LLMs) has been empirically effective on a variety of downstream tasks. Existing approaches to finetuning an LLM either focus on parameter-efficient finetuning, which only updates a small number of trainable parameters, or attempt to reduce the memory footprint during the training phase of the finetuning. Typically, the memory footprint during finetuning stems from

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).