← all papers · overview

Openfactcheck: Building, Benchmarking Customized Fact-checking Systems And Evaluating The Factuality Of Claims And Llms

Abstract

The increased use of large language models (LLMs) across a variety of real-world applications calls for mechanisms to verify the factual accuracy of their outputs. Difficulties lie in assessing the factuality of free-form responses in open domains. Also, different papers use disparate evaluation benchmarks and measurements, which renders them hard to compare and hampers future progress. To mitigat

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).