← all papers · overview

Importance Weighting Can Help Large Language Models Self-improve

Abstract

Large language models (LLMs) have shown remarkable capability in numerous tasks and applications. However, fine-tuning LLMs using high-quality datasets under external supervision remains prohibitively expensive. In response, LLM self-improvement approaches have been vibrantly developed recently. The typical paradigm of LLM self-improvement involves training LLM on self-generated data, part of whic

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).