← all papers · overview

Allora: Adaptive Learning Rate Mitigates Lora Fatal Flaws

Abstract

Low-Rank Adaptation (LoRA) is the bread and butter of Large Language Model (LLM) finetuning. LoRA learns an additive low-rank perturbation, , of a pretrained matrix parameter to align the model to a new task or dataset with . We identify three core limitations to LoRA for finetuning--a setting that employs limited amount of data and training steps. First, LoRA employs Dropout t

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).