datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ASR_Fellowship_Challenge_DatasetAfrivoice_Swahili_0.0
Dataset summary
Domain
Total number of hours
Total number of transcribed hours
Total number of clips
Total Size of the dataset in GB
Agriculture
35.19
35.19
5,852
2.6
Health
62.11
62.11
10,315
6.4
Finance
118.75
118.75
19,629
16.6
Government
106.41
106.41
17,530
12.1
Education
91.96
91.96
15,204
6.2
Total
414.42
414.42
68,530
43.9
How to use
The datasets library allows you to load and pre-process your dataset in pure Python, at scale. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/DigitalUmuganda/Afrivoice_Swahili_0.0.Afrivoice_Kinyarwanda_old_version
Dataset summary
[need more information]
Supported tasks
[need more information]
How to use
[need more information]
Dataset structure
Data fields
[need more information]
Data splits
[need more information]
Data preprocessing
[need more information]
Licensing Information
All datasets are licensed under the Creative Commons license (CC-BY-4).
