CoolFace
Datasetpublic

Professor/kinyarwanda-speech-data

Kinyarwanda Speech Data (Pooled) A ~989.9-hour pooled Kinyarwanda speech corpus, combining two independently-sourced datasets into one consistently-formatted corpus for speech modeling (TTS / ASR). Part of the AfroNet multi-language TTS data effort — sibling release to the Yoruba/Hausa/Igbo pools, but sourced entirely differently: DSN African Voices, NaijaVoices, and WAXAL (the sources behind the other three languages) don't cover Kinyarwanda at all. Sources… See the full description on the dataset page: https://huggingface.co/datasets/Professor/kinyarwanda-speech-data.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
0likes14downloads
filemanifest.parquet22.9 MBdownload

Professor/kinyarwanda-speech-data · main · files are served by the source, never re-hosted here