← all papers · overview

Structext-eval: Evaluating Large Language Model's Reasoning Ability In Structure-rich Text

Abstract

The effective utilization of structured data, integral to corporate data strategies, has been challenged by the rise of large language models (LLMs) capable of processing unstructured information. This shift prompts the question: can LLMs interpret structured data directly in its unstructured form? We propose an automatic evaluation data generation method for assessing LLMs' reasoning capabilities

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).