← all papers · overview

Know When You're Wrong: Aligning Confidence With Correctness For LLM Error Detection

Abstract

As large language models (LLMs) are increasingly deployed in critical decision-making systems, the lack of reliable methods to measure their uncertainty presents a fundamental trustworthiness risk. We introduce a normalized confidence score based on output anchor token probabilities: classification labels for structured tasks and self-evaluation responses (Yes/No) for open-ended generation. This e

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).