← all papers · overview

LLM Internal States Reveal Hallucination Risk Faced With A Query

Abstract

The hallucination problem of Large Language Models (LLMs) significantly limits their reliability and trustworthiness. Humans have a self-awareness process that allows us to recognize what we don't know when faced with queries. Inspired by this, our paper investigates whether LLMs can estimate their own hallucination risk before response generation. We analyze the internal mechanisms of LLMs broadl

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).