← all papers · overview

High-fidelity Pruning For Large Language Models

Abstract

Large Language Models (LLMs) have demonstrated exceptional performance across a wide range of tasks, yet their significant computational and memory requirements present major challenges for deployment. A common approach uses Taylor expansion on the loss function to estimate neuron importance. However, its reliance on one-hot cross entropy loss, a key limitation is that it narrowly assesses importa

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).