← all papers · overview

Beyond The Black Box: A Statistical Model For LLM Reasoning And Inference

Abstract

This paper introduces a novel Bayesian learning model to explain the behavior of Large Language Models (LLMs), focusing on their core optimization metric of next token prediction. We develop a theoretical framework based on an ideal generative text model represented by a multinomial transition probability matrix with a prior, and examine how LLMs approximate this matrix. Key contributions include:

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).