CoolFace
Datasetpublicgated

naijavoices/voiceafrica-datasets

VoiceAfrica Dataset Introduction Welcome to the VoiceAfrica dataset. VoiceAfrica is a Lanfrica–Meta collaboration (code-named VoiceAfrica 1) that set out to create authentic, conversational speech and expert-curated transcriptions for 11 under-represented African languages. The dataset contains ~125 hours of speech (about 10 hours per language) across 16,383 audio samples from 118 speakers in four countries. Unlike read-speech corpora, VoiceAfrica uses a natural… See the full description on the dataset page: https://huggingface.co/datasets/naijavoices/voiceafrica-datasets.

sourceHugging Facecc-by-nc-sa-4.0updated 1mo agoView on Hugging Face
3likes25downloads

naijavoices/voiceafrica-datasets · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.