← all papers · overview

Evaluating Zero-shot Long-context LLM Compression

Abstract

This study evaluates the effectiveness of zero-shot compression techniques on large language models (LLMs) under long-context. We identify the tendency for computational errors to increase under long-context when employing certain compression methods. We propose a hypothesis to explain the varied behavior of different LLM compression techniques and explore remedies to mitigate the performance decl

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).