← all papers · overview

"vorbeşti Româneşte?" A Recipe To Train Powerful Romanian Llms With English Instructions

Abstract

In recent years, Large Language Models (LLMs) have achieved almost human-like performance on various tasks. While some LLMs have been trained on multilingual data, most of the training data is in English; hence, their performance in English greatly exceeds other languages. To our knowledge, we are the first to collect and translate a large collection of texts, instructions, and benchmarks and trai

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).