Professor/dholuo-speech-data
Dholuo Speech Data (Pooled) A ~191.5-hour Dholuo (Luo) speech corpus, pooled from two independent sources and filtered to only genuinely transcribed audio. Part of the AfroNet multi-language TTS data effort. Sources Anv-ke/Dholuo — African Next Voices, a pilot data-collection effort in Kenya led by the KenCorpus Consortium (a coalition of Kenyan universities and research centers), funded by the Gates Foundation. 91,672 clips, 186.1h, source = anv_ke. Gated on… See the full description on the dataset page: https://huggingface.co/datasets/Professor/dholuo-speech-data.
Add dataset card
Add audio shards
Add manifest.jsonl
Add manifest.parquet
Add dataset card
Add audio shards
Add manifest.jsonl
Add manifest.parquet
initial commit
