← all papers · overview

Genaudit: Fixing Factual Errors In Language Model Outputs With Evidence

Abstract

LLMs can generate factually incorrect statements even when provided access to reference documents. Such errors can be dangerous in high-stakes applications (e.g., document-grounded QA for healthcare or finance). We present GenAudit -- a tool intended to assist fact-checking LLM responses for document-grounded tasks. GenAudit suggests edits to the LLM response by revising or removing claims that ar

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).