DigitalUmuganda/Afrivoice
Dataset Card for the image text and voice dataset Dataset Description Each datapoint in this dataset consists of a JPEG image, a corresponding audio WAV file describing the image, and when available, the transcription of the audio file. Language Language Code Total Audio Hours Transcribed Audio Hours Shona sn 574.16 100.00 Lingala ln 517.13 100.98 Fulani ful 527.45 102.00 Malagasy mg 516.21 102.42 Wolof wo 530.74 102.96 Somali so 535.51… See the full description on the dataset page: https://huggingface.co/datasets/DigitalUmuganda/Afrivoice.
1326
No card is published for this repository, or it could not be fetched from Hugging Face right now.
