← all papers · overview

Diff-erank: A Novel Rank-based Metric For Evaluating Large Language Models

Abstract

Large Language Models (LLMs) have transformed natural language processing and extended their powerful capabilities to multi-modal domains. As LLMs continue to advance, it is crucial to develop diverse and appropriate metrics for their evaluation. In this paper, we introduce a novel rank-based metric, Diff-eRank, grounded in information theory and geometry principles. Diff-eRank assesses LLMs by an

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).