← all papers · overview

Automated Data Curation For Robust Language Model Fine-tuning

Abstract

Large Language Models have become the de facto approach to sequence-to-sequence text generation tasks, but for specialized tasks/domains, a pretrained LLM lacks specific capabilities to produce accurate or well-formatted responses. Supervised fine-tuning specializes a LLM by training it on dataset of example prompts with target responses, but real-world data tends to be noisy. While many fine-tuni

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).