← all datasets

LongBench

Emerging
6papers using it
2024first seen

LongBench is a comprehensive benchmark for multilingual and multi-task purposes, with the goal to fully measure and evaluate the ability of pre-trained language models to understand long text. This dataset consists of twenty different tasks, covering key long-text application scenarios such as multi-document QA, single

Papers using LongBench (6)

LongBench dataset β€” papers, benchmarks & downloads Β· AI Agents