← all papers · overview

Efficient Deployment Of Large Language Models On Resource-constrained Devices

Abstract

Deploying Large Language Models (LLMs) on resource-constrained (or weak) devices presents significant challenges due to limited resources and heterogeneous data distribution. To address the data concern, it is necessary to fine-tune LLMs using on-device private data for various downstream tasks. While Federated Learning (FL) offers a promising privacy-preserving solution, existing fine-tuning meth

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).