← all papers · overview

Bypass Back-propagation: Optimization-based Structural Pruning For Large Language Models Via Policy Gradient

Abstract

Recent Large-Language Models (LLMs) pruning methods typically operate at the post-training phase without the expensive weight finetuning, however, their pruning criteria often rely on heuristically hand-crafted metrics, potentially leading to suboptimal performance. We instead propose a novel optimization-based structural pruning that learns the pruning masks in a probabilistic space directly by o

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).