← all papers · overview

Performance Of Small Language Model Pretraining On FABRIC: An Empirical Study

Abstract

Large language models (LLMs) require enormous computing power to pretrain on massive datasets. When limited datasets are available, smaller-sized LLMs are better choice to pretrain (on user-specified datasets) by following the scaling laws of LLMs. Using pretrained models, vector embeddings can be generated for raw data and stored using vector databases to support modern AI applications and semant

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).