← all papers · overview

Blockllm: Memory-efficient Adaptation Of Llms By Selecting And Optimizing The Right Coordinate Blocks

Abstract

Training large language models (LLMs) for pretraining or adapting to new tasks and domains has become increasingly critical as their applications expand. However, as the model and the data sizes grow, the training process presents significant memory challenges, often requiring a prohibitive amount of GPU memory that may not be readily available. Existing methods such as low-rank adaptation (LoRA)

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).