← all papers · overview

Towards A Holistic Evaluation Of Llms On Factual Knowledge Recall

Abstract

Large language models (LLMs) have shown remarkable performance on a variety of NLP tasks, and are being rapidly adopted in a wide range of use cases. It is therefore of vital importance to holistically evaluate the factuality of their generated outputs, as hallucinations remain a challenging issue. In this work, we focus on assessing LLMs' ability to recall factual knowledge learned from pretrai

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).