← all datasets

WikiText-2

Emerging
1papers using it
2025first seen

WikiText-2 is a benchmark dataset that contains a collection of over 2 million tokens extracted from Wikipedia articles, used to evaluate the performance of language models in terms of their ability to generate coherent and contextually relevant text.

Papers using WikiText-2 (1)

WikiText-2 dataset β€” papers, benchmarks & downloads Β· AI Agents