← all papers · overview

Dissecting The Runtime Performance Of The Training, Fine-tuning, And Inference Of Large Language Models

Abstract

Large Language Models (LLMs) have seen great advance in both academia and industry, and their popularity results in numerous open-source frameworks and techniques in accelerating LLM pre-training, fine-tuning, and inference. Training and deploying LLMs are expensive as it requires considerable computing resources and memory, hence many efficient approaches have been developed for improving system

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).