← all papers · overview

Testing And Understanding Erroneous Planning In LLM Agents Through Synthesized User Inputs

Abstract

Agents based on large language models (LLMs) have demonstrated effectiveness in solving a wide range of tasks by integrating LLMs with key modules such as planning, memory, and tool usage. Increasingly, customers are adopting LLM agents across a variety of commercial applications critical to reliability, including support for mental well-being, chemical synthesis, and software development. Neverth

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).