← all papers · overview

Dynamic Gradient Alignment For Online Data Mixing

Abstract

The composition of training data mixtures is critical for effectively training large language models (LLMs), as it directly impacts their performance on downstream tasks. Our goal is to identify an optimal data mixture to specialize an LLM for a specific task with access to only a few examples. Traditional approaches to this problem include ad-hoc reweighting methods, importance sampling, and grad

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).