← all papers · overview

Language Model Planners Do Not Scale, But Do Formalizers?

Abstract

Recent work shows overwhelming evidence that LLMs, even those trained to scale their reasoning trace, perform unsatisfactorily when solving planning problems too complex. Whether the same conclusion holds for LLM formalizers that generate solver-oriented programs remains unknown. We systematically show that LLM formalizers greatly out-scale LLM planners, some retaining perfect accuracy in the clas

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).