← all papers · overview

Empirical Guidelines For Deploying Llms Onto Resource-constrained Edge Devices

Abstract

The scaling laws have become the de facto guidelines for designing large language models (LLMs), but they were studied under the assumption of unlimited computing resources for both training and inference. As LLMs are increasingly used as personalized intelligent assistants, their customization (i.e., learning through fine-tuning) and deployment onto resource-constrained edge devices will become m

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).