← all papers · overview

Achieving Sparse Activation In Small Language Models

Abstract

Sparse activation, which selectively activates only an input-dependent set of neurons in inference, is a useful technique to reduce the computing cost of Large Language Models (LLMs) without retraining or adaptation efforts. However, whether it can be applied to the recently emerging Small Language Models (SLMs) remains questionable, because SLMs are generally less over-parameterized than LLMs. In

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).