datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
shoebox_rir_with_room_namestamil_names_audiotamil_names_audio_v2chemical-namessingaporean_accent_district_names_dataset_ultravoxsingaporean_accent_district_names_continuationname_synthesizedpolymer-namesafri-names
Afri-names: Read Speech Dataset of Numbers and African Named Entities
This work is licensed under aCreative Commons Attribution-NonCommercial-ShareAlike 4.0 International License.
Overview
Afri-names is a curated African-accented read speech dataset comprising 6,307 single-speaker audio samples, totaling 8.92 hours of speech data. Each sample is densely populated with numbers or African named entities or voice commands (with African named entities), making it ideal… See the full description on the dataset page: https://huggingface.co/datasets/intronhealth/afri-names.test-large-names_speaker_similaritytest-base-names_speaker_similarityembedded_world_2026_rag_tts_INC_names
