← all papers · overview

Sparse Models, Sparse Safety: Unsafe Routes In Mixture-of-experts Llms

Abstract

By introducing routers to selectively activate experts in Transformer layers, the mixture-of-experts (MoE) architecture significantly reduces computational costs in large language models (LLMs) while maintaining competitive performance, especially for models with massive parameters. However, prior work has largely focused on utility and efficiency, leaving the safety risks associated with this spa

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).