← all papers · overview

Uncovering Hidden Correctness In LLM Causal Reasoning Via Symbolic Verification

Abstract

Large language models (LLMs) are increasingly being applied to tasks that involve causal reasoning. However, current benchmarks often rely on string matching or surface-level metrics that do not capture whether the output of a model is formally valid under the semantics of causal reasoning. To address this, we propose DoVerifier, a simple symbolic verifier that checks whether LLM-generated causal

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).