← all papers · overview

Strategist: Self-improvement Of LLM Decision Making Via Bi-level Tree Search

Abstract

Traditional reinforcement learning and planning typically requires vast amounts of data and training to develop effective policies. In contrast, large language models (LLMs) exhibit strong generalization and zero-shot capabilities, but struggle with tasks that require detailed planning and decision-making in complex action spaces. We introduce STRATEGIST, a novel approach that integrates the stren

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).