← all papers · overview

Xlstm 7B: A Recurrent LLM For Fast And Efficient Inference

Abstract

Recent breakthroughs in solving reasoning, math and coding problems with Large Language Models (LLMs) have been enabled by investing substantial computation budgets at inference time. Therefore, inference speed is one of the most critical properties of LLM architectures, and there is a growing need for LLMs that are efficient and fast at inference. Recently, LLMs built on the xLSTM architecture ha

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).