← all papers · overview

Vinallama: Llama-based Vietnamese Foundation Model

Abstract

In this technical report, we present VinaLLaMA, an open-weight, state-of-the-art (SOTA) Large Language Model for the Vietnamese language, built upon LLaMA-2 with an additional 800 billion trained tokens. VinaLLaMA not only demonstrates fluency in Vietnamese but also exhibits a profound understanding of Vietnamese culture, making it a truly indigenous model. VinaLLaMA-7B-chat, trained on 1 million

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).