← all papers · overview

Computational Bottlenecks Of Training Small-scale Large Language Models

Abstract

While large language models (LLMs) dominate the AI landscape, Small-scale large Language Models (SLMs) are gaining attention due to cost and efficiency demands from consumers. However, there is limited research on the training behavior and computational requirements of SLMs. In this study, we explore the computational bottlenecks of training SLMs (up to 2B parameters) by examining the effects of v

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).