← all papers · overview

The Veln(ia)s Is In The Details: Evaluating LLM Judgment On Latvian And Lithuanian Short Answer Matching

Abstract

In this work, we address the challenge of evaluating large language models (LLMs) on the short answer matching task for Latvian and Lithuanian languages. We introduce novel datasets consisting of 502 Latvian and 690 Lithuanian question-answer pairs. For each question-answer pair, we generated matched and non-matched answers using a set of alteration rules specifically designed to introduce small b

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).