← all datasets

LJSpeech

Canonical
37papers using it
2021first seen

This is a public domain speech dataset consisting of 13,100 short audio clips of a single speaker reading passages from 7 non-fiction books. A transcription is provided for each clip. Clips vary in length from 1 to 10 seconds and have a total length of approximately 24 hours.

Papers using LJSpeech (37)

LJSpeech dataset — papers, benchmarks & downloads · Speech Audio