← all papers · overview

Constructing Benchmarks And Interventions For Combating Hallucinations In Llms

Abstract

Large language models (LLMs) are prone to hallucinations, which sparked a widespread effort to detect and prevent them. Recent work attempts to mitigate hallucinations by intervening in the model's generation, typically computing representative vectors of hallucinations vs. grounded generations, for steering the model's hidden states away from a hallucinatory state. However, common studies employ

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).