← all papers · overview

Gradmap: Faster Layer Pruning With Gradient Metric And Projection Compensation

Abstract

Large Language Models (LLMs) exhibit strong reasoning abilities, but their high computational costs limit their practical deployment. Recent studies reveal significant redundancy in LLMs layers, making layer pruning an active research topic. Layer pruning research primarily focuses on two aspects: measuring layer importance and recovering performance after pruning. Unfortunately, the present works

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).