CoolFace
Datasetpublic

ghananlpcommunity/twi-tts-asr

Ghana Twi Speech Dataset (TTS + ASR) 79,655 synthetic Twi speech samples (88.8 hours) generated with OmniVoice. Designed for both text-to-speech (TTS) and automatic speech recognition (ASR) research on Twi. Intended use TTS / Voice cloning: Use the audio + text pairs directly. The dataset includes diverse voice profiles for building or fine-tuning Twi TTS systems. ASR: Use the text transcripts as labels. The audio is natively at 24 kHz — resample to 16 kHz for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/twi-tts-asr.

sourceHugging Facecc-by-nc-4.0updated 2mo agoView on Hugging Face
0likes449downloads

ghananlpcommunity/twi-tts-asr · main · files are served by the source, never re-hosted here