CoolFace
Datasetpublic

ahmed220v/mgb2-arabic

MGB-2: Arabic Multi-Dialect Broadcast Media Recognition Dataset Description Dataset Summary The Arabic Multi-Genre Broadcast (MGB-2) dataset is a large-scale speech recognition corpus containing 1,200 hours of Arabic broadcast audio from Aljazeera Arabic TV channel. The dataset spans recordings from March 2005 to December 2015 and covers 19 distinct programme series. It was originally created for the MGB-2 Challenge at SLT-2016, focusing on handling… See the full description on the dataset page: https://huggingface.co/datasets/ahmed220v/mgb2-arabic.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes274downloads
settings

This repository belongs to ahmed220v on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namemgb2-arabic
visibilitypublic
licencenot set
gatedno
ownerahmed220v
Account settings
ahmed220v/mgb2-arabic · CoolFace