← all papers · overview

Basis Selection: Low-rank Decomposition Of Pretrained Large Language Models For Target Applications

Abstract

Large language models (LLMs) significantly enhance the performance of various applications, but they are computationally intensive and energy-demanding. This makes it challenging to deploy them on devices with limited resources, such as personal computers and mobile/wearable devices, and results in substantial inference costs in resource-rich environments like cloud servers. To extend the use of L

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).