← all papers · overview

Hyperbolic Fine-tuning For Large Language Models

Abstract

Large language models (LLMs) have demonstrated remarkable performance across various tasks. However, it remains an open question whether the default Euclidean space is the most suitable choice for LLMs. In this study, we investigate the geometric characteristics of LLMs, focusing specifically on tokens and their embeddings. Our findings reveal that token frequency follows a power-law distribution,

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).