← all papers · overview

BAS: A Decision-theoretic Approach To Evaluating Large Language Model Confidence

Abstract

Large language models (LLMs) often produce confident but incorrect answers in settings where abstention would be safer. Standard evaluation protocols, however, require a response and do not account for how confidence should guide decisions under different risk preferences. To address this gap, we introduce the Behavioral Alignment Score (BAS), a decision-theoretic metric for evaluating how well LL

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).