CoolFace
Datasetpublic

MohamedGomaa30/Egyptian-Speech-Clean-MGB3

🏛️ Dataset Card for MGB3-Egyptian-Clean Dataset Summary This dataset is a refined and enhanced version of the MGB-3 (Multi-Genre Broadcast) corpus, specifically focused on the Egyptian Arabic dialect. It has been meticulously preprocessed to be "TTS-ready" by combining advanced deep-learning denoising with custom linguistic text normalization. 🛠️ Preprocessing Pipeline To ensure the highest quality for generative speech tasks (like VITS or MMS… See the full description on the dataset page: https://huggingface.co/datasets/MohamedGomaa30/Egyptian-Speech-Clean-MGB3.

sourceHugging Faceupdated 8mo agoView on Hugging Face
0likes52downloads
settings

This repository belongs to MohamedGomaa30 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameEgyptian-Speech-Clean-MGB3
visibilitypublic
licencenot set
gatedno
ownerMohamedGomaa30
Account settings
MohamedGomaa30/Egyptian-Speech-Clean-MGB3 · CoolFace