CoolFace
Datasetpublic

ymoslem/SpokenWords-GA-EN-MTed

Dataset Card for Dataset Name This is the Irish portion of the Spoken Words dataset (available at MLCommons/ml_spoken_words), with merged splits “train”, “validation”, and “test”, augmented with machine translation. The Irish sentences are automatically translated into English using Google Translation API. The dataset includes approximately 3 hours and 2 minutes of audio (03:02:02), spoken by multiple narrators. Dataset Structure Dataset({ features: ['keyword'… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/SpokenWords-GA-EN-MTed.

sourceHugging Facemitupdated 2y agoView on Hugging Face
1likes68downloads
7 commits on main
b8c4ded2y ago

Update README.md

ymoslem
983fe432y ago

Update README.md

ymoslem
4616a732y ago

Update README.md

ymoslem
de70c752y ago

Upload dataset

ymoslem
9d4da0e2y ago

Update README.md

ymoslem
679396f2y ago

Upload dataset

ymoslem
deee1cb2y ago

initial commit

ymoslem