← all papers · overview

Adapt-pruner: Adaptive Structural Pruning For Efficient Small Language Model Training

Abstract

Small language models (SLMs) have attracted considerable attention from both academia and industry due to their broad range of applications in edge devices. To obtain SLMs with strong performance, conventional approaches either pre-train the models from scratch, which incurs substantial computational costs, or compress/prune existing large language models (LLMs), which results in performance drops

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).