CoolFace
Datasetpublic

ymoslem/Tatoeba-Speech-Irish

Dataset Details Synthetic audio dataset, created using Azure text-to-speech service. The bilingual text is a portion of the Tatoeba dataset, consisting of 1,983 text segments. The dataset consists of two sets of audio data, one with a female voice (OrlaNeural) and the other with a male voice (ColmNeural). The speech data comprises approximately 2 hours and 39 minutes (02:39:31) spread across 3,966 utterances. Dataset Structure Dataset({ features: ['audio'… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/Tatoeba-Speech-Irish.

sourceHugging Facecc-by-4.0updated 2y agoView on Hugging Face
1likes72downloads
10 commits on main
daae62c2y ago

Update README.md

ymoslem
c3360592y ago

Update README.md

ymoslem
4cb83272y ago

Update README.md

ymoslem
22d9d172y ago

Update README.md

ymoslem
39c10d72y ago

Update README.md

ymoslem
72952df2y ago

Update README.md

ymoslem
4210ccd2y ago

Update README.md

ymoslem
6d355ac2y ago

Upload dataset

ymoslem
cf7d6a32y ago

Upload dataset

ymoslem
7e919bb2y ago

initial commit

ymoslem