← all papers · overview

Ranking Llms By Compression

Abstract

We conceptualize the process of understanding as information compression, and propose a method for ranking large language models (LLMs) based on lossless data compression. We demonstrate the equivalence of compression length under arithmetic coding with cumulative negative log probabilities when using a large language model as a prior, that is, the pre-training phase of the model is essentially th

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).