← all papers · overview

Breaking Language Barriers: Cross-lingual Continual Pre-training At Scale

Abstract

In recent years, Large Language Models (LLMs) have made significant strides towards Artificial General Intelligence. However, training these models from scratch requires substantial computational resources and vast amounts of text data. In this paper, we explore an alternative approach to constructing an LLM for a new language by continually pretraining (CPT) from existing pretrained LLMs, instead

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).