← all papers · overview

Transferable Text Data Distillation By Trajectory Matching

Abstract

In the realm of large language model (LLM), as the size of large models increases, it also brings higher training costs. There is a urgent need to minimize the data size in LLM training. Compared with data selection method, the data distillation method aims to synthesize a small number of data samples to achieve the training effect of the full data set and has better flexibility. Despite its succe

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).