← all papers · overview

Understanding LLM Behaviors Via Compression: Data Generation, Knowledge Acquisition And Scaling Laws

Abstract

Large Language Models (LLMs) have demonstrated remarkable capabilities across numerous tasks, yet principled explanations for their underlying mechanisms and several phenomena, such as scaling laws, hallucinations, and related behaviors, remain elusive. In this work, we revisit the classical relationship between compression and prediction, grounded in Kolmogorov complexity and Shannon information

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).