← all papers · overview

Dna-eval: Enhancing Large Language Model Evaluation Through Decomposition And Aggregation

Abstract

The acceleration of Large Language Models (LLMs) research has opened up new possibilities for evaluating generated texts. They serve as scalable and economical evaluators, but the question of how reliable these evaluators are has emerged as a crucial research question. Prior research efforts in the meta-evaluation of LLMs as judges limit the prompting of an LLM to a single use to obtain a final ev

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).