← all papers · overview

Joyai-llm Flash: Advancing Mid-scale Llms With Token Efficiency

Abstract

We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performance and token efficiency in the sub-50B parameter regime. JoyAI-LLM Flash is pretrained on a massive corpus of 20 trillion tokens and further optimized through a rigorous post-training pipeline, including supervised fine-tuning (SFT), Direct Preference Optimi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).