← all papers · overview

Geometric Analysis Of Token Selection In Multi-head Attention

Abstract

We present a geometric framework for analysing multi-head attention in large language models (LLMs). Without altering the mechanism, we view standard attention through a top-N selection lens and study its behaviour directly in value-state space. We define geometric metrics - Precision, Recall, and F-score - to quantify separability between selected and non-selected tokens, and derive non-asymptoti

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).