← all papers · overview

Data-free Weight Compress And Denoise For Large Language Models

Abstract

Large Language Models (LLMs) are reshaping the research landscape in artificial intelligence, particularly as model parameters scale up significantly, unlocking remarkable capabilities across various domains. Nevertheless, the scalability of model parameters faces constraints due to limitations in GPU memory and computational speed. To address these constraints, various weight compression methods

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).