CoolFace
Datasetpublicgated

DigitalUmuganda/Afrivoice

Dataset Card for the image text and voice dataset Dataset Description Each datapoint in this dataset consists of a JPEG image, a corresponding audio WAV file describing the image, and when available, the transcription of the audio file. Language Language Code Total Audio Hours Transcribed Audio Hours Shona sn 574.16 100.00 Lingala ln 517.13 100.98 Fulani ful 527.45 102.00 Malagasy mg 516.21 102.42 Wolof wo 530.74 102.96 Somali so 535.51… See the full description on the dataset page: https://huggingface.co/datasets/DigitalUmuganda/Afrivoice.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
1likes326downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.