← all papers · overview

Nanoknow: How To Know What Your Language Model Knows

Abstract

How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a "black box" -- unknown or inaccessible. The recent release of nanochat -- a family of small LLMs with fully open pre-training data -- addresses this as it provides a transparent view into where a model's parametric knowledge comes from. Towards the goal of unders

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).