← all papers · overview

When Silence Is Golden: Can Llms Learn To Abstain In Temporal QA And Beyond?

Abstract

Large language models (LLMs) rarely admit uncertainty, often producing fluent but misleading answers, rather than abstaining (i.e., refusing to answer). This weakness is even evident in temporal question answering, where models frequently ignore time-sensitive evidence and conflate facts across different time-periods. In this paper, we present the first empirical study of training LLMs with an abs

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).