← all papers · overview

Seedlm: Compressing LLM Weights Into Seeds Of Pseudo-random Generators

Abstract

Large Language Models (LLMs) have transformed natural language processing, but face significant challenges in widespread deployment due to their high runtime cost. In this paper, we introduce SeedLM, a novel post-training compression method that uses seeds of pseudo-random generators to encode and compress model weights. Specifically, for each block of weights, we find a seed that is fed into a Li

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).