← all papers · overview

Hyperparameter Optimization For Large Language Model Instruction-tuning

Abstract

The fine-tuning of Large Language Models (LLMs) has enabled them to recently achieve milestones in natural language processing applications. The emergence of ever larger LLMs has paved the way for more efficient fine-tuning methods. Among these, the Low-Rank Adaptation (LoRA) method keeps most of the weights of the pre-trained LLM frozen while introducing a low-rank decomposition of the weight mat

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).