← all papers · overview

Mods: Model-oriented Data Selection For Instruction Tuning

Abstract

Instruction tuning has become the de facto method to equip large language models (LLMs) with the ability of following user instructions. Usually, hundreds of thousands or millions of instruction-following pairs are employed to fine-tune the foundation LLMs. Recently, some studies show that a small number of high-quality instruction data is enough. However, how to select appropriate instruction dat

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).