← all papers · overview

On The Self-verification Limitations Of Large Language Models On Reasoning And Planning Tasks

Abstract

There has been considerable divergence of opinion on the reasoning abilities of Large Language Models (LLMs). While the initial optimism that reasoning might emerge automatically with scale has been tempered thanks to a slew of counterexamples--ranging from multiplication to simple planning--there persists a wide spread belief that LLMs can self-critique and improve their own solutions in an itera

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).