← all papers · overview

Evaluating Interventional Reasoning Capabilities Of Large Language Models

Abstract

Numerous decision-making tasks require estimating causal effects under interventions on different parts of a system. As practitioners consider using large language models (LLMs) to automate decisions, studying their causal reasoning capabilities becomes crucial. A recent line of work evaluates LLMs ability to retrieve commonsense causal facts, but these evaluations do not sufficiently assess how L

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).