← all datasets

1.6M books

Emerging
1papers using it
2020first seen

The '1.6M books' dataset is a benchmark containing a large collection of text from books used to evaluate the performance of embedding models in terms of query efficiency and accuracy.

Papers using 1.6M books (1)

1.6M books dataset β€” papers, benchmarks & downloads Β· Similarity Search