← all papers · overview

AC-EVAL: Evaluating Ancient Chinese Language Understanding In Large Language Models

Abstract

Given the importance of ancient Chinese in capturing the essence of rich historical and cultural heritage, the rapid advancements in Large Language Models (LLMs) necessitate benchmarks that can effectively evaluate their understanding of ancient contexts. To meet this need, we present AC-EVAL, an innovative benchmark designed to assess the advanced knowledge and reasoning capabilities of LLMs with

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).