← all papers · overview

Foundabench: Evaluating Chinese Fundamental Knowledge Capabilities Of Large Language Models

Abstract

In the burgeoning field of large language models (LLMs), the assessment of fundamental knowledge remains a critical challenge, particularly for models tailored to Chinese language and culture. This paper introduces FoundaBench, a pioneering benchmark designed to rigorously evaluate the fundamental knowledge capabilities of Chinese LLMs. FoundaBench encompasses a diverse array of 3354 multiple-choi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).