CoolFace
Datasetpublic

facebook/multilingual_librispeech

Dataset Card for MultiLingual LibriSpeech Dataset Summary This is a streamable version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish.… See the full description on the dataset page: https://huggingface.co/datasets/facebook/multilingual_librispeech.

sourceHugging Facecc-by-4.0updated 2y agoView on Hugging Face
190likes34kdownloads
settings

This repository belongs to facebook on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namemultilingual_librispeech
visibilitypublic
licencecc-by-4.0
gatedno
ownerfacebook
Account settings
facebook/multilingual_librispeech · CoolFace