← all papers · overview

Automated Factual Benchmarking For In-car Conversational Systems Using Large Language Models

Abstract

In-car conversational systems bring the promise to improve the in-vehicle user experience. Modern conversational systems are based on Large Language Models (LLMs), which makes them prone to errors such as hallucinations, i.e., inaccurate, fictitious, and therefore factually incorrect information. In this paper, we present an LLM-based methodology for the automatic factual benchmarking of in-car co

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).