← all papers · overview

Self-moe: Towards Compositional Large Language Models With Self-specialized Experts

Abstract

We present Self-MoE, an approach that transforms a monolithic LLM into a compositional, modular system of self-specialized experts, named MiXSE (MiXture of Self-specialized Experts). Our approach leverages self-specialization, which constructs expert modules using self-generated synthetic data, each equipping a shared base LLM with distinct domain-specific capabilities, activated via self-optimize

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).