CoolFace
Datasetpublic

ehabnegm/100-hour-Egyptian-dataset-single-speaker

Masri 100h — Egyptian Arabic Single-Speaker Speech Corpus A 100-hour Egyptian Arabic (مصري) single-narrator speech collection — 15,653 released clips at 24 kHz mono, with aligned transcripts. Egyptian Arabic is the most widely understood Arabic dialect and one of the least served by open speech data. Almost every open Arabic corpus is Modern Standard Arabic (MSA) — a register nobody actually speaks at home. This dataset is built for the opposite: natural, spoken, conversational… See the full description on the dataset page: https://huggingface.co/datasets/ehabnegm/100-hour-Egyptian-dataset-single-speaker.

sourceHugging Facecc-by-nc-4.0updated 2mo agoView on Hugging Face
11likes4.9kdownloads
settings

This repository belongs to ehabnegm on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

name100-hour-Egyptian-dataset-single-speaker
visibilitypublic
licencecc-by-nc-4.0
gatedno
ownerehabnegm
Account settings
ehabnegm/100-hour-Egyptian-dataset-single-speaker · CoolFace