← all papers · overview

Visualizing And Benchmarking LLM Factual Hallucination Tendencies Via Internal State Analysis And Clustering

Abstract

Large Language Models (LLMs) often hallucinate, generating nonsensical or false information that can be especially harmful in sensitive fields such as medicine or law. To study this phenomenon systematically, we introduce FalseCite, a curated dataset designed to capture and benchmark hallucinated responses induced by misleading or fabricated citations. Running GPT-4o-mini, Falcon-7B, and Mistral 7

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).