CoolFace
Datasetpublicgated

naijavoices/voiceafrica-datasets

VoiceAfrica Dataset Introduction Welcome to the VoiceAfrica dataset. VoiceAfrica is a Lanfrica–Meta collaboration (code-named VoiceAfrica 1) that set out to create authentic, conversational speech and expert-curated transcriptions for 11 under-represented African languages. The dataset contains ~125 hours of speech (about 10 hours per language) across 16,383 audio samples from 118 speakers in four countries. Unlike read-speech corpora, VoiceAfrica uses a natural… See the full description on the dataset page: https://huggingface.co/datasets/naijavoices/voiceafrica-datasets.

sourceHugging Facecc-by-nc-sa-4.0updated 1mo agoView on Hugging Face
3likes25downloads

No commit history came back for main. The revision may not exist, or the source declined the request.