WikiText-2
Emerging1papers using it
2025first seen
WikiText-2 is a benchmark dataset that contains a collection of over 2 million tokens extracted from Wikipedia articles, used to evaluate the performance of language models in terms of their ability to generate coherent and contextually relevant text.