← all papers · overview

Direct Evaluation Of Chain-of-thought In Multi-hop Reasoning With Knowledge Graphs

Abstract

Large language models (LLMs) demonstrate strong reasoning abilities when prompted to generate chain-of-thought (CoT) explanations alongside answers. However, previous research on evaluating LLMs has solely focused on answer accuracy, neglecting the correctness of the generated CoT. In this paper, we delve deeper into the CoT reasoning capabilities of LLMs in multi-hop question answering by utilizi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).