← all papers · overview

Llmeval: A Preliminary Study On How To Evaluate Large Language Models

Abstract

Recently, the evaluation of Large Language Models has emerged as a popular area of research. The three crucial questions for LLM evaluation are ``what, where, and how to evaluate''. However, the existing research mainly focuses on the first two questions, which are basically what tasks to give the LLM during testing and what kind of knowledge it should deal with. As for the third question, which i

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).