← all papers · overview

On Evaluating The Integration Of Reasoning And Action In LLM Agents With Database Question Answering

Abstract

This study introduces a new long-form database question answering dataset designed to evaluate how Large Language Models (LLMs) interact with a SQL interpreter. The task necessitates LLMs to strategically generate multiple SQL queries to retrieve sufficient data from a database, to reason with the acquired context, and to synthesize them into a comprehensive analytical narrative. Our findings high

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).