CoolFace
Datasetpublic

Professor/kinyarwanda-speech-data

Kinyarwanda Speech Data (Pooled) A ~989.9-hour pooled Kinyarwanda speech corpus, combining two independently-sourced datasets into one consistently-formatted corpus for speech modeling (TTS / ASR). Part of the AfroNet multi-language TTS data effort — sibling release to the Yoruba/Hausa/Igbo pools, but sourced entirely differently: DSN African Voices, NaijaVoices, and WAXAL (the sources behind the other three languages) don't cover Kinyarwanda at all. Sources… See the full description on the dataset page: https://huggingface.co/datasets/Professor/kinyarwanda-speech-data.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
0likes14downloads
5 commits on main
da500b21mo ago

Add dataset card

Professor
d9a9b041mo ago

Add audio shards

Professor
3ff6eaf1mo ago

Add manifest.jsonl

Professor
8fe2a6f1mo ago

Add manifest.parquet

Professor
0e5644a1mo ago

initial commit

Professor