← all papers · overview

52B To 1T: Lessons Learned Via Tele-flm Series

Abstract

Large Language Models (LLMs) represent a significant stride toward Artificial General Intelligence. As scaling laws underscore the potential of increasing model sizes, the academic community has intensified its investigations into LLMs with capacities exceeding 50 billion parameters. This technical report builds on our prior work with Tele-FLM (also known as FLM-2), a publicly available 52-billion

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).