ymoslem/Tatoeba-Speech-Irish
Dataset Details Synthetic audio dataset, created using Azure text-to-speech service. The bilingual text is a portion of the Tatoeba dataset, consisting of 1,983 text segments. The dataset consists of two sets of audio data, one with a female voice (OrlaNeural) and the other with a male voice (ColmNeural). The speech data comprises approximately 2 hours and 39 minutes (02:39:31) spread across 3,966 utterances. Dataset Structure Dataset({ features: ['audio'… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/Tatoeba-Speech-Irish.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face