← all papers · overview

Elitr-bench: A Meeting Assistant Benchmark For Long-context Language Models

Abstract

Research on Large Language Models (LLMs) has recently witnessed an increasing interest in extending the models' context size to better capture dependencies within long documents. While benchmarks have been proposed to assess long-range abilities, existing efforts primarily considered generic tasks that are not necessarily aligned with real-world applications. In contrast, we propose a new benchmar

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).