← all papers · overview

Self-consistency Of Large Language Models Under Ambiguity

Abstract

Large language models (LLMs) that do not give consistent answers across contexts are problematic when used for tasks with expectations of consistency, e.g., question-answering, explanations, etc. Our work presents an evaluation benchmark for self-consistency in cases of under-specification where two or more answers can be correct. We conduct a series of behavioral experiments on the OpenAI model s

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).