← all papers · overview

Moreaupruner: Robust Pruning Of Large Language Models Against Weight Perturbations

Abstract

Few-shot gradient methods have been extensively utilized in existing model pruning methods, where the model weights are regarded as static values and the effects of potential weight perturbations are not considered. However, the widely used large language models (LLMs) have several billion model parameters, which could increase the fragility of few-shot gradient pruning. In this work, we experimen

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).