← all papers · overview

Representing The Under-represented: Cultural And Core Capability Benchmarks For Developing Thai Large Language Models

Abstract

The rapid advancement of large language models (LLMs) has highlighted the need for robust evaluation frameworks that assess their core capabilities, such as reasoning, knowledge, and commonsense, leading to the inception of certain widely-used benchmark suites such as the H6 benchmark. However, these benchmark suites are primarily built for the English language, and there exists a lack thereof for

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).