← all papers · overview

Instructdiff: Domain-adaptive Data Selection Via Differential Entropy For Efficient LLM Fine-tuning

Abstract

Supervised fine-tuning (SFT) is fundamental to adapting large language models, yet training on complete datasets incurs prohibitive costs with diminishing returns. Existing data selection methods suffer from severe domain specificity: techniques optimized for general instruction-following fail on reasoning tasks, and vice versa. We observe that measuring entropy differences between base models and

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).