← all papers · overview

Two-stage Optimizer-aware Online Data Selection For Large Language Models

Abstract

Gradient-based data selection offers a principled framework for estimating sample utility in large language model (LLM) fine-tuning, but existing methods are mostly designed for offline settings. They are therefore less suited to online fine-tuning, where data arrives sequentially, sample utility is step-dependent, and the effective update geometry is shaped by adaptive optimizers. We propose an o

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).