← all papers · overview

Uncertainty-based Abstention In Llms Improves Safety And Reduces Hallucinations

Abstract

A major barrier towards the practical deployment of large language models (LLMs) is their lack of reliability. Three situations where this is particularly apparent are correctness, hallucinations when given unanswerable questions, and safety. In all three cases, models should ideally abstain from responding, much like humans, whose ability to understand uncertainty makes us refrain from answering

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).