← all papers · overview

Omgeval: An Open Multilingual Generative Evaluation Benchmark For Large Language Models

Abstract

Modern large language models (LLMs) should generally benefit individuals from various cultural backgrounds around the world. However, most recent advanced generative evaluation benchmarks tailed for LLMs mainly focus on English. To this end, we introduce OMGEval, the first Open-source Multilingual Generative test set that can assess the capability of LLMs in different languages. For each language,

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).