← all papers · overview

Logicgame: Benchmarking Rule-based Reasoning Abilities Of Large Language Models

Abstract

Large Language Models (LLMs) have demonstrated notable capabilities across various tasks, showcasing complex problem-solving abilities. Understanding and executing complex rules, along with multi-step planning, are fundamental to logical reasoning and critical for practical LLM agents and decision-making systems. However, evaluating LLMs as effective rule-based executors and planners remains under

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).