datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ghana-named-entities-tts-twi
This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/.
Ghana Named Entities TTS — Twi
A Twi-language speech dataset built from descriptions of Ghana named entities
(people, places, organisations, and concepts). Each audio clip is a synthesised
reading of a passage that describes several… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-named-entities-tts-twi.afri-names
Afri-names: Read Speech Dataset of Numbers and African Named Entities
This work is licensed under aCreative Commons Attribution-NonCommercial-ShareAlike 4.0 International License.
Overview
Afri-names is a curated African-accented read speech dataset comprising 6,307 single-speaker audio samples, totaling 8.92 hours of speech data. Each sample is densely populated with numbers or African named entities or voice commands (with African named entities), making it ideal… See the full description on the dataset page: https://huggingface.co/datasets/intronhealth/afri-names.
