legacy-datasets/multilingual_librispeech
Multilingual LibriSpeech (MLS) dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish.
17181
This repository belongs to legacy-datasets on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
multilingual_librispeech
public
cc-by-4.0
no
legacy-datasets
