← all papers · overview

Look Further Ahead: Testing The Limits Of GPT-4 In Path Planning

Abstract

Large Language Models (LLMs) have shown impressive capabilities across a wide variety of tasks. However, they still face challenges with long-horizon planning. To study this, we propose path planning tasks as a platform to evaluate LLMs' ability to navigate long trajectories under geometric constraints. Our proposed benchmark systematically tests path-planning skills in complex settings. Using thi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).