CoolFace
Datasetpublic

LEMAS-Project/LEMAS-Dataset-eval

Overview This dataset is part of LEMAS-Project(lemas-project.github.io/LEMAS-Project). It contains a large-scale training set (150k+ hours) and a curated evaluation set (500 utterances per language) covering 10 languages, all with word-level alignment. Fields key: unique utterance identifier; the first two characters indicate the language ID audio: relative path to the MP3 audio file (in the eval set, this key is renamed to "file_name" for compatibility with the… See the full description on the dataset page: https://huggingface.co/datasets/LEMAS-Project/LEMAS-Dataset-eval.

sourceHugging Facecc-by-4.0updated 6mo agoView on Hugging Face
2likes111downloads
9 commits on main
538d9af6mo ago

Update README.md

Approximetal
0e996909mo ago

Update README.md

Approximetal
9cb9aeb9mo ago

Update README.md

Approximetal
65d7c959mo ago

Update README.md

Approximetal
06505a99mo ago

Update README.md

Approximetal
a2c82149mo ago

Update README.md

Approximetal
3b75f819mo ago

Create README.md

Approximetal
253a4369mo ago

Upload folder using huggingface_hub

Approximetal
0a13c3f9mo ago

initial commit

Approximetal