← all papers · overview

Mind The Gap: Examining The Self-improvement Capabilities Of Large Language Models

Abstract

Self-improvement is a mechanism in Large Language Model (LLM) pre-training, post-training and test-time inference. We explore a framework where the model verifies its own outputs, filters or reweights data based on this verification, and distills the filtered data. Despite several empirical successes, a fundamental understanding is still lacking. In this work, we initiate a comprehensive, modular

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).