← all papers · overview

Walking Down The Memory Maze: Beyond Context Limit Through Interactive Reading

Abstract

Large language models (LLMs) have advanced in large strides due to the effectiveness of the self-attention mechanism that processes and compares all tokens at once. However, this mechanism comes with a fundamental issue -- the predetermined context window is bound to be limited. Despite attempts to extend the context window through methods like extrapolating the positional embedding, using recurre

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).