← all papers · overview

Understanding Large Language Model Behaviors Through Interactive Counterfactual Generation And Analysis

Abstract

Understanding the behavior of large language models (LLMs) is crucial for ensuring their safe and reliable use. However, existing explainable AI (XAI) methods for LLMs primarily rely on word-level explanations, which are often computationally inefficient and misaligned with human reasoning processes. Moreover, these methods often treat explanation as a one-time output, overlooking its inherently i

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).