← all papers · overview

Zero-order Optimization For LLM Fine-tuning Via Learnable Direction Sampling

Abstract

Fine-tuning large pretrained language models (LLMs) is a cornerstone of modern NLP, yet its growing memory demands (driven by backpropagation and large optimizer States) limit deployment in resource-constrained settings. Zero-order (ZO) methods bypass backpropagation by estimating directional derivatives from forward evaluations, offering substantial memory savings. However, classical ZO estimator

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).