← all papers · overview

Vcllm: Video Codecs Are Secretly Tensor Codecs

Abstract

As the parameter size of large language models (LLMs) continues to expand, the need for a large memory footprint and high communication bandwidth have become significant bottlenecks for the training and inference of LLMs. To mitigate these bottlenecks, various tensor compression techniques have been proposed to reduce the data size, thereby alleviating memory requirements and communication pressur

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).