← all papers · overview

Pruning Via Merging: Compressing Llms Via Manifold Alignment Based Layer Merging

Abstract

While large language models (LLMs) excel in many domains, their complexity and scale challenge deployment in resource-limited environments. Current compression techniques, such as parameter pruning, often fail to effectively utilize the knowledge from pruned parameters. To address these challenges, we propose Manifold-Based Knowledge Alignment and Layer Merging Compression (MKA), a novel approach

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).