← all papers · overview

Weak-to-strong Reasoning

Abstract

When large language models (LLMs) exceed human-level capabilities, it becomes increasingly challenging to provide full-scale and accurate supervision for these models. Weak-to-strong learning, which leverages a less capable model to unlock the latent abilities of a stronger model, proves valuable in this context. Yet, the efficacy of this approach for complex reasoning tasks is still untested. Fur

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).