← all papers · overview

Large Language Models Must Be Taught To Know What They Don't Know

Abstract

When using large language models (LLMs) in high-stakes applications, we need to know when we can trust their predictions. Some works argue that prompting high-performance LLMs is sufficient to produce calibrated uncertainties, while others introduce sampling methods that can be prohibitively expensive. In this work, we first argue that prompting on its own is insufficient to achieve good calibrati

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).