← all papers · overview

Assessing Logical Puzzle Solving In Large Language Models: Insights From A Minesweeper Case Study

Abstract

Large Language Models (LLMs) have shown remarkable proficiency in language understanding and have been successfully applied to a variety of real-world tasks through task-specific fine-tuning or prompt engineering. Despite these advancements, it remains an open question whether LLMs are fundamentally capable of reasoning and planning, or if they primarily rely on recalling and synthesizing informat

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).