CoolFace
Datasetpublicgated

ivrit-ai/crowd-whatsapp-yi

About This dataset was created by crowd-sourced Whatsapp voice recordings in Yiddish as part of the ivrit.ai project. Volunteers read a message sent to them from a predefined set of messages, recording themselves using Whasapp voice message sent to the collecting bot. Later this data is normalized by aligning the captions with the audio using Stable Whisper (See Below). The recording project is an ongoing effort and new data will be appended to this dataset periodically as it is… See the full description on the dataset page: https://huggingface.co/datasets/ivrit-ai/crowd-whatsapp-yi.

sourceHugging Faceotherupdated 10mo agoView on Hugging Face
0likes24downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.

ivrit-ai/crowd-whatsapp-yi · CoolFace