← all papers · overview

Knowledge-based Consistency Testing Of Large Language Models

Abstract

In this work, we systematically expose and measure the inconsistency and knowledge gaps of Large Language Models (LLMs). Specifically, we propose an automated testing framework (called KonTest) which leverages a knowledge graph to construct test cases. KonTest probes and measures the inconsistencies in the LLM's knowledge of the world via a combination of semantically-equivalent queries and test o

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).