← all papers · overview

Explain-query-test: Self-evaluating Llms Via Explanation And Comprehension Discrepancy

Abstract

Large language models (LLMs) have demonstrated remarkable proficiency in generating detailed and coherent explanations of complex concepts. However, the extent to which these models truly comprehend the concepts they articulate remains unclear. To assess the level of comprehension of a model relative to the content it generates, we implemented a self-evaluation pipeline where models: (i) given a t

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).