← all papers · overview

AA-SVD : Anchored And Adaptive SVD For Large Language Model Compression

Abstract

We introduce a fast low-rank factorization-based framework for compressing large language models that enables rapid compression of billion-parameter models without retraining. Unlike existing factorization-based approaches that optimize only on the original inputs, ignoring distribution shifts from upstream compression and thus propagating errors forward, or those that rely only on shifted inputs

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).