← all papers · overview

Pruning Large Language Models Via Accuracy Predictor

Abstract

Large language models(LLMs) containing tens of billions of parameters (or even more) have demonstrated impressive capabilities in various NLP tasks. However, substantial model size poses challenges to training, inference, and deployment so that it is necessary to compress the model. At present, most model compression for LLMs requires manual design of pruning features, which has problems such as c

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).