← all papers · overview

Disentangling Deception And Hallucination Failures In Llms

Abstract

Failures in large language models (LLMs) are often analyzed from a behavioral perspective, where incorrect outputs in factual question answering are commonly associated with missing knowledge. In this work, focusing on entity-based factual queries, we suggest that such a view may conflate different failure mechanisms, and propose an internal, mechanism-oriented perspective that separates Knowledge

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).