← all papers · overview

Defan: Definitive Answer Dataset For Llms Hallucination Evaluation

Abstract

Large Language Models (LLMs) have demonstrated remarkable capabilities, revolutionizing the integration of AI in daily life applications. However, they are prone to hallucinations, generating claims that contradict established facts, deviating from prompts, and producing inconsistent responses when the same prompt is presented multiple times. Addressing these issues is challenging due to the lack

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).