← all papers · overview

Accelerating Large Language Model Pretraining Via LFR Pedagogy: Learn, Focus, And Review

Abstract

Traditional Large Language Model (LLM) pretraining relies on autoregressive language modeling with randomly sampled data from web-scale datasets. Inspired by human learning techniques like spaced repetition, we hypothesize that random sampling leads to high training costs, lower-quality models, and significant data forgetting. To address these inefficiencies, we propose the Learn-Focus-Review (LFR

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).