← all papers · overview

Reconfidencing Llms From The Grouping Loss Perspective

Abstract

Large Language Models (LLMs), including ChatGPT and LLaMA, are susceptible to generating hallucinated answers in a confident tone. While efforts to elicit and calibrate confidence scores have proven useful, recent findings show that controlling uncertainty must go beyond calibration: predicted scores may deviate significantly from the actual posterior probabilities due to the impact of grouping lo

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).