← all papers · overview

Can Large Language Models Learn Independent Causal Mechanisms?

Abstract

Despite impressive performance on language modelling and complex reasoning tasks, Large Language Models (LLMs) fall short on the same tasks in uncommon settings or with distribution shifts, exhibiting a lack of generalisation ability. By contrast, systems such as causal models, that learn abstract variables and causal relationships, can demonstrate increased robustness against changes in the distr

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).