← all papers · overview

Get More For Less: Principled Data Selection For Warming Up Fine-tuning In Llms

Abstract

This work focuses on leveraging and selecting from vast, unlabeled, open data to pre-fine-tune a pre-trained language model. The goal is to minimize the need for costly domain-specific data for subsequent fine-tuning while achieving desired performance levels. While many data selection algorithms have been designed for small-scale applications, rendering them unsuitable for our context, some emerg

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).