← all papers · overview

Zeroth-order Fine-tuning Of Llms In Random Subspaces

Abstract

Fine-tuning Large Language Models (LLMs) has proven effective for a variety of downstream tasks. However, as LLMs grow in size, the memory demands for backpropagation become increasingly prohibitive. Zeroth-order (ZO) optimization methods offer a memory-efficient alternative by using forward passes to estimate gradients, but the variance of gradient estimates typically scales linearly with the mod

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).