← all papers · overview

From Tokens To Agents: A Researcher's Guide To Understanding Large Language Models

Abstract

Researchers face a critical choice: how to use -- or not use -- large language models in their work. Using them well requires understanding the mechanisms that shape what LLMs can and cannot do. This chapter makes LLMs comprehensible without requiring technical expertise, breaking down six essential components: pre-training data, tokenization and embeddings, transformer architecture, probabilistic

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).