← all papers · overview

Gradient Weight-normalized Low-rank Projection For Efficient LLM Training

Abstract

Large Language Models (LLMs) have shown remarkable performance across various tasks, but the escalating demands on computational resources pose significant challenges, particularly in the extensive utilization of full fine-tuning for downstream tasks. To address this, parameter-efficient fine-tuning (PEFT) methods have been developed, but they often underperform compared to full fine-tuning and st

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).