← all papers · overview

Giraffe: Adventures In Expanding Context Lengths In Llms

Abstract

Modern large language models (LLMs) that rely on attention mechanisms are typically trained with fixed context lengths which enforce upper limits on the length of input sequences that they can handle at evaluation time. To use these models on sequences longer than the train-time context length, one might employ techniques from the growing family of context length extrapolation methods -- most of w

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).