← all papers · overview

Evaluating Llm-based Translation Of A Low-resource Technical Language: The Medical And Philosophical Greek Of Galen

Abstract

Purpose: This study evaluates the quality of commercial large language model (LLM) machine translation (MT) for Ancient Greek technical prose and benchmarks standard automated MT evaluation metrics against expert human judgment. Design: We evaluated 60 translations by three LLMs (ChatGPT, Claude, Gemini) of 20 paragraph-length passages from 2

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).