← all papers · overview

SPAFIT: Stratified Progressive Adaptation Fine-tuning For Pre-trained Large Language Models

Abstract

Full fine-tuning is a popular approach to adapt Transformer-based pre-trained large language models to a specific downstream task. However, the substantial requirements for computational power and storage have discouraged its widespread use. Moreover, increasing evidence of catastrophic forgetting and overparameterization in the Transformer architecture has motivated researchers to seek more effic

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).