← all papers · overview

PEMA: An Offsite-tunable Plug-in External Memory Adaptation For Language Models

Abstract

Pre-trained language models (PLMs) show impressive performance in various downstream NLP tasks. However, pre-training large language models demands substantial memory and training compute. Furthermore, due to the substantial resources required, many PLM weights are confidential. Consequently, users are compelled to share their data with model owners for fine-tuning specific tasks. To overcome the

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).