← all datasets

Multilingual LibriSpeech

Canonical
12papers using it
2021first seen

Dataset Card for MultiLingual LibriSpeech Dataset Summary This is a streamable version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset

Papers using Multilingual LibriSpeech (12)

Multilingual LibriSpeech dataset β€” papers, benchmarks & downloads Β· Speech Audio