← all papers · overview

The Future Of Large Language Model Pre-training Is Federated

Abstract

Generative pre-trained large language models (LLMs) have demonstrated impressive performance over a wide range of tasks, thanks to the unprecedented amount of data they have been trained on. As established scaling laws indicate, LLMs' future performance improvement depends on the amount of computing and data sources they can leverage for pre-training. Federated learning (FL) has the potential to u

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).