← all papers · overview

A Theoretical Understanding Of Self-correction Through In-context Alignment

Abstract

Going beyond mimicking limited human experiences, recent studies show initial evidence that, like humans, large language models (LLMs) are capable of improving their abilities purely by self-correction, i.e., correcting previous responses through self-examination, in certain circumstances. Nevertheless, little is known about how such capabilities arise. In this work, based on a simplified setup ak

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).